a ; ^'-^^^ 10/067,800 

(12) INTERNATIONAt APPLICATION PUBUSHED UNDER THE PATENT COOPERATION TREATY (PCT) 



(19) World Intellectual Property Organization 
Intematibridl.Bureaii 




liiilliiiliiiiiiliU 



(43) International Publication Date (10) International Publication Number 

20 December 2001 (20.12.2001) pCT WO 01/96388 A2 



(51) International Patent Classification^: C07K 14/47 



(21) International Application Number: PCT/US01/1S557 



(22) International Filing Date: 8 June 2001 (08.06.2001) 



(25) Filing Language: English 

(26) Pablication Language: English 

(30) Priority Data: 

60/210,899 9 June 2000 (09.06.2000) US 

60^270^16 20 February 2001 (20J)2.2001) US 



(71) Applicant (^or all designated States except US): CORIXA 
CORPORATION [USAJS]; Suite 200, 1124 Columbia 
Street. Seattle, WA 98104 (US). 



s=5 (72) Inventors; and 

(72) Inventors/Applicants (for US only): JIANG, Yuqin 
= [CNAJSJ; 5001 a 232nd Street, Kent. WA 98032 (US). 
= HARLOCKER, Susan, L. [USAJS]; 7522 13th Av- 
= enue W., Seattle. WA 98117 (US). SfiCRIST, Heather 
= [USAJS]; 3844 35th Avenue W., Seattle, WA 98199 (US). 



(74) Agents: POTTER, ^gne, R-; Seed Intellectual Prop- 
erty Law Group PLLC, Suite 6300, 701 Fifth Avenue, Seat- 
tle, WA 98 104-7092. et al. (US). 

(81) Designated States (national): AE, AG, AU AM, AT, AU, 
AZ, BA, BB, BG, BR, BY. BZ, CA, CH, CN, CO, CR, CU. 
CZ. DB, DK, DM. DZ, EC, EE, ES, FI, OB, GD. GE, GH, 
GM, HR, HU, ID, IL, IN, IS, JP, KE, KG. KP, KR, KZ, LC. 
LK, LR, LS, LT, LU, LV, MA, MD, MG, MK, MN, MW, 
MX, MZ, NO, NZ, PL, FT. RO. RU, SD, SE, SG, SI, SK. 
SL, TJ, TM, TR, TT. TZ. UA, UG, US, UZ, VN, YU, ZA, 
ZW. 

(84) Designated States (regional): ARIPO patent (GH, GM, 
KE, LS, MW, MZ, SD, SL, SZ, TZ, UG. ZW), Eurasian 
patent (AM, AZ. BY, KG, KZ, MD, RU, TJ . TM), European 
patent (AT. BE. CH. C Y, DE, DK, ES, FI, FR, GB, GR, IE, 
rr, LU, MC, NL, PT. SE, TR), OAK patent (BF, BJ, CF, 
CG. O. CM , G A. GN, GW, ML. MR, NE, SN, TO, TO). 

Published: 

— without international search r^ort and to be republished 
upon receipt cf that report 

For two-letter codes and other abbreviations, refer to the **Guid' 
ance Notes on Codes and Abbreviations" appearing at the begin- 
nir^ of each regular issue cf the PCT Gazette. 



00 

00 

rO (54) rule: COMPOSITIONS AND METHODS FOR THE THERAPY AND DUGNOSIS OF COLON CANCER 

^ (57) Abstract: Compositions and methods for the therapy and diagnosis of cancer, such as colon cancer, are disclosed. Composi- 
^ tions may comprise one or more colon tumor proteins, immunogenic portions thereof, or polynucleotides that encode such poitions. 

Alternatively, a therapeutic composition may comprise an antigen presenting cell that expresses a colon tumor protein, or a T cell 
^ that is specific for cells expressing such a protein. Such compositions may be used, for example, for the prevention and treatment of 
^ diseases such as colon cancer. Diagnostic methods based on detecting a colon tumor protein, or mRNA encoding such a protein, in 
1^ a sample are also provided. 



BNSDOaD: <WO___0t963e8A^L> 



BEST AVAILABLE COPY 



wo 01796388 PCT/US01/18SS7 



COMPOSITIONS AND METHODS FOR THE THERAPY AND DIAGNOSIS 

OF COLON CANCER 

TECHNICAL FIELD OF THE INVENTION _ 

The. present invention relates generally to therapy and diagnosis of 
5 cancer, such as colon cancer. The invention is more specificaUy related to polypeptides 
comprising at least a portion of a colon tumor protein, and to polynucleotides encoding 
such polypeptides. Such polypeptides and polynucleotides may be used in vaccines and 
pharmaceutical compositions for prevention and treatment of colon malignancies, and 
for the diagnosis and monitoring of such cancers. 

10 BACKGROUND OF THE INVENTION 

Cancer is a significant health problem fhrougjiout the world. Although 
advances have been made in detection and therapy of cancer, no vaccine or other 
universally successful method for prevention or treatment is currently available. 
Current therapies, which are generally based on a combination of chemothoapy or 

1 S surgery and radiation, continue to prove inadequate in many patients. 

Colon cancer is the second most frequently diagnosed malignancy in the 
United States as well as the second most common cause of cancer death. The five-year 
survival rate for patients with colorectal cancer detected in an early localized stage is 
92%; unfortunately, only 37% of colorectal cancer, is diagnosed at this stage. The 

20 survival rate drops to 64% if the cancer is allowed to spread to adjacent organs or 
lymph nodes, and to 7% in patients with distant metastases. 

The prognosis of colon canc^ is directly related to the degree of 
penetration of the tumor through the bowel wall and the presence or absence of nodal 
involvemrat, consequentiy early detection and treatment are eq>ecially important 

25 Currently, diagnosis is aided by the use of screening assays for fecal occult blood, 
sigmoidoscopy, colonoscopy and double contrast barium enemas. Treatment regimens 
are determined by the type and stage of the cancer, and include surgery, radiation 
therapy and/or chemotherapy. Recurrence following surgery (the most common form 
of therapy) is a major problem and is often the ultimate cause of death. 



BNSOOCIO: <WO__OI96388A?JL> 



wo 01/96388 



2 



PCTAJSOl/18557 



In spite of considerable research into therapies for these and other 
cancers, colon cancer remains difficult to diagnose and treat effectively. Accordingly, 
there is a need in the art for improved methods for detecting and treating such cancers. 
The present invention fulfills these needs and fiirth^ provides other related advantages. 

5 SUMMARY OF THE INVENTION 

In one aspect, the present invention provides polynucleotide 
compositions comprising a sequrace selected fix>m the group consisting of: 

(a) sequences provided in SEQ ID NO:l-2234; 

(b) complements of the sequences provided in SEQ ID NO:l-2234; 
10 (c) sequences consisting of at least 20, 25, 30, 35, 40, 45, 50, 75 and 

100 contiguous residues of a sequence provided in SEQ ID NO: 1-2234; 

(d) sequences that hybridize to a sequence provided in SEQ ID 
NO: 1-2234, under moderate or highly stringent conditions; 

(e) sequences havmg at least 75%, 80%, 85%, 90%, 95%, 96%, 
15 97%, 98% or 99% identity to a sequence of SEQ ID NO:l-2234; 

(f) degenerate variants of a sequence provided in SEQ ID NO:l- 

2234. 

In one preferred embodiment, the polynucleotide compositions of the 
invention are expressed in at least about 20%, more preferably in at least about 30%, 
20 and most preferably m at least about 50% of colon tumor samples tested, at a level that 
is at least about 2*fold, preferably at least about S-fold, and most preferably at least 
about 10-fold higher than that for normal tissues. 

The present invention, in another aspect, provides polypeptide 
compositions comprising an amino acid sequence that is encoded by a polynucleotide 
25 sequence described above. 

The present invention further provides polypeptide compositions 
comprising an amino acid sequence selected from the groi^ consisting of the sequence 
recited in SEQ ID NO:2235. 

In certain preferred embodiments, the polypeptides and/or 
30 polynucleotides of the present invention are immunogenic, ie., they are capable of 



BNSOOCID: <W0 ^01963B8AaJL^ 



wo 01/96388 



3 



PCTAJSOl/18557 



eliciting an immune response, particularly a humoral and/or cellular immune response, 
as further described herein. 

The preset invention further provides fragments, variants and/or 
derivatives of the disclosed polypeptide and/or polynucleotide sciences, wherein the 
5 fragments^ variants and/or derivatives preferably have a level of immunogenic activity 
of at least about 50%, preferably at least about 70% and more preferably at least about 
90% of the level of immunogenic activity of a polypq)tide sequence set forth in SEQ 
ID NO:2235 or a polypeptide sequence encoded by a polynucleotide sequence set forth 
in SEQ ID NO: 1-2234. 

10 The present invention further provides polynucleotides that encode a 

polypeptide described above, expression vectors comprising such polynucleotides and 

host cells transfonned or transfected with such expression vectors. 

Within other aspects, the present invention provides pharmaceutical 

compositions comprising a polypeptide or polynucleotide as described above and a 

1 5 physiologically acceptable carrier. 

Withm a related aspect of the present invention, the pharmaceutical 

compositions, e.g., vaccine compositions, are provided for prophylactic or therapeutic 

applications. Such compositions generally comprise an immunogenic polypeptide or 

polynucleotide of the invention and an immunostimuiant, such as an adjuvant 
20 The present invention furtiber provides pharmaceutical compositions that 

comprise: (a) an antibody or antigen-binding fragment thereof that specifically bmds to 

a polypeptide of the present invention, or a fiagment thereof; and (b) a physiologically 

acceptable carrier. * 

Within fiirther aspects, the present invention provides pharmaceutical 

25 compositions comprising: (a) an antigen presenting ceU that expresses a polypq>tide as 

described above and (b) a pharmaceutically acceptable carrier or excipient Illustrative 

antigen presenting cells include dendritic cells, macrophages, monocytes, fibroblasts 

and B cells* 

Within related aspects, pharmaceutical compositions are provided that 
30 comprise: (a) an antigen presenting cell that expresses a polypeptide as described 
above and (b) an immunostimuiant. 



BNSDOCII>:_<WO ^96388A9JL> 



wo 01/96388 PCTAJSOl/18557 

4 

The present invention further provides, in other aspects, fusion proteins 
that comprise at least one polypeptide as described above, as well as polynucleotides 
encoding such fusion proteins, typically in the form of pharmaceutical compositions, 
e.g., vaccine compositions, comprising a physiologically acceptable canier and/or an 
5 immunostimulant The fusions proteins may comprise multiple immunogenic 
polypeptides or portions/variants thereof, as described herein, and may further comprise 
one or more polypeptide segments for facilitating the expression, purification and/or 
immunogenicity of the polypeptide(s). 

Within further aspects, the present mvention provides methods for 
10 stimulating an immune response in a patient, preferably a T cell response in a human 
patient, comprismg administering a pharmaceutical composition described herein. The 
patient may be afQicted with colon cancer, in which case the methods provide treatment 
for tiie disease, or patient considered at risk for such a disease may be treated 
prophylactically. 

15 Within further aspects, the present invention provides methods for 

inhibiting the development of a cancer in a patient, comprising administering to a 
patient a pharmaceutical conq)Osition as recited above. The patirat may be afQicted 
with colon cancer, in which case the methods provide treatment for the disease, or 
patient considered at risk for such a disease may be treated prophylactically. 

20 The present invention further provides, within other aspects, methods for 

removing tumor cells from a biological sample, comprising contacting a biological 
sample with T cells that specifically react with a polypq>tide of the present invention, 
wherein the step of contactmg is performed under conditions and for a time sufiBcient to 
permit the removal of cells expressmg the protein from the sample. 

25 Within related aspects, methods are provided for inhibiting the 

development of a cancer m a patient, comprising administering to a patient a biological 
sample treated as described above. 

Methods are further provided, within other aspects, for stimulating 
and/or expanding T cells specific for a polypeptide of the present invention, comprising 

30 contacting T ceUs with one or more of: (i) a polypeptide as described above; (ii) a 
polynucleotide encoding such a polypeptide; and/or (iii) an antigen presenting cell that 
expresses such a polypeptide; under conditions and for a time sufficient to permit the 



BNSOOCID: 



_01963B8A2_L> 



wo 01/96388 PCT/USOl/18557 

5 

stimulation and/or expansion of T cells. Isolated T cell populations comprising T ceUs 
prepared as described above are also provided. 

Within further aspects, the present invention provides methods for 
inhibiting the development of a cancer in a patient, comprising administering to a 
5 patient an effective amount of a T cell population as described above. 

The present invention further provides methods for inhibiting the 
development of a cancer in a patient, comprising the steps of: (a) incubating CD4+ 
and/or CDS'*" T cells isolated from a patient with one or more of: Q) a polypeptide 
comprising at least an immunogenic portion of polypeptide disclosed herein; (ii) a 
10 polynucleotide encoding such a polypeptide; and (iii) an antigen-presenting cell that 
expressed such a polypeptide; and (b) administering to the patient an effective amount 
of the proliferated T cells, and thereby inhibiting the develoinn^ of a cancer m the 
patient Proliferated cells may, but need not, be cloned prior to administration to the 
, patient. 

15 Within further aspects, the present invention provides methods for 

determining the presence or absence of a cancer, preferably a colon cancer, in a patient 
comprising: (a) contacting a biological sample obtained from a patient with a binding 
agent that binds to a polypeptide as recited above; (b) detecting in the sample an 
amount of polypeptide that binds to the bmding agent; and (c) comparing the amount of 

20 polypeptide with a predetermined cut-off value, and therefrom determining the presence 
or absence of a cancer in the patient. Within preferred embodiments, the binding agent 
is an antibody, more preferably a monoclonal antibody. 

The present invention also provides, within other aspects, methods for 
monitoring the progression of a cancer in a patient. Such methods conqprise the steps 

25 of: (a) contacting a biological sample obtained from a patient at a first point m time 
with a binding agent that binds to a polypeptide as recited above; (b) detecting in the 
sample an amount of polypeptide that binds to the bmding agent; (c) repeating steps (a) 
and (b) using a biological sample obtained fix>m die patirat at a subsequent point in 
time; and (d) comparing the amount of polypeptide detected in step (c) with the amount 

30 detected in step (b) and therefrom monitoring the progression of the cancer in the 
. patient. 



BNSDOCIO: <W0 ^019638aASUL> 



wo 01/96388 PCT/US01/J8557 

6 

The present invention fiirflier provides, within otiier aspects, methods for 
determining the presence or absence of a cancer in a patient, comprising the steps of: (a) 
contacting a biological sample, e.g., tumor sample, serum sample, etc., obtained from a 
patient with an oligonucleotide that hybridizes to a polynucleotide that encodes a 

5 polypeptide of the present mvention; (b) detecting in the -sample a level of a 
polynucleotide, preferably mRNA, that hybridizes to the oligonucleotide; and (c) 
comparing the level of polynucleotide that hybridizes to the oligonucleotide with a 
predetermined cut-off value, and therefrom determining the presence or absence of a 
cancer in the patient. Within certain embodiments, the amount of mRNA is detected 

10 via polymerase chain reaction using, for example, at least one oligonucleotide primer 
that hybridizes to a polynucleotide encoding a polypeptide as recited above, or a 
complement of such a polynucleotide. Within other embodiments, the amount of 
mRNA is detected uang a hybridization technique, employing an oligonucleotide probe 
that hybridizes to a polynucleotide that encodes a polypeptide as recited above, or a 

1 S complement of such a polynucleotide. 

In related aspects, methods are provided for monitoring the progression 
of a cancer in a patient, comprismg the steps of: (a) contacting a biological sample 
obtained from a patient with an oligonucleotide that hybridizes to a polynucleotide tihat 
encodes a polypq)tide of the present mvention; (b) detecting in the sample an amount of 

20 a polynucleotide that hybridizes to the oligonucleotide; (c) repeating steps (a) and (b) 
using a biolo^cal sample obtained from the patient at a subsequent point m time; and 
(d) comparmg the amount of polynucleotide detected in step (c) with the amount 
detected in step (b) and therefrom monitoring the progression of the cancer in the 
patient. 

25 * Within further aspects, the present invention provides antibodies, such as 

monoclonal antibodies, that bind to a polypeptide as described above, as well as 
diagnostic kits comprising such antibodies. Diagnostic kits compriang one or more 
oligonucleotide probes or primers as described above are also provided. 

These and other aspects of the present invention will become apparent 

30 upon reference to the following detailed description and attached. All references 
disclosed herein are hereby incorporated by reference m their entirety as if each was 
incorporated individually. 



BNSOOaD: <WO___01Be388A?JL? 



wo Dl/%388 



PCT/USOl/18557 



7 

DETAILED DESCRIPTION OF THE INVENTION 

The present invention is directed generally to compositions and their use 
in the fterapy and diagnosis of cancer, particularly colon cancer. As described further 
below, illustrative compositions of the present invention include J>ut are not restricted 
5 to, polypeptides, particulariy immunogenic polypeptides, polynucleotides encoding 
such polypeptides, antibodies and other binding agents, antigen presenting cells (APCs) 
and immune system cells (e,g.j T cells). 

The practice of the present invention >vill employ, unless indicated 
specifically to the contrary, conventional methods of virology, immunology, 

10 microbiology, molecular biology and recombinant DNA techniques within the skill of 
the art, many of which are described below for the purpose of illustratian. Such 
techniques are explained fully in the lit^nfoire. See, e.g., Sambrook, et al. Molecular 
Cloning: A Laboratory Manual (2nd Edition, 1989); Maniatis et al* Molecular Clonmg: 
A Laboratory Manual (1982); DNA Cloning: A Practical Approach, vol. I & II (D. 

15 Glover, ei); Oligonucleotide Synthesis (N. Gait, ed., 1984); Nucleic Add 
Hybridization (B. Hames & S. Higgins, eds., 1985); Transcription and Translation (B. 
Hames & S. Higgins, eds., 1984); Animal Cell Culture (R. Freshney, ed., 1986); Perbal, 
A Practical Guide to Molecular Cloning (1984). 

All publications, patents and patent applications cited herein, whether 

20 supra or infra, are hereby incorporated by reference in their entirety. 

As used in Hds specification and the appended claims, the lingular forms 
''a,'* "an" and "the" include plural references unless the content clearly dictates 
otherwise. 

POLYPEFTIDE COMPOSITIONS 

25 As used herein, the term "polypeptide" " is used in its conventional 

meaning, z.e., as a sequence of amino acids. The polypeptides are not limited to a 
specific length of the product; thus, peptides, oligopeptides, and proteins are included 
within the definition of polypeptide, and such terms may be used mterchangeably 
herein unless specifically indicated otherwise. This term also does not refer to or 

30 exclude post-expression modifications of the polypeptide, for example, glycosylations. 



wo 01/96388 



8 



PCT/USOl/18557 



acetylations, phosphorylations and the like, as well as other modifications known in the 
art, both naturally occurring and non-naturally occurring. A polypeptide may be an 
ratire protein, or a subsequence fliereof. Particular polypeptides of interest in Ae 
context of this invention are amino acid subsequences coniprising epitopes, /.e., 

5 antigenic deteiminants substantially responsible for the inamwiogenic properties of a 
polypeptide and being capable of evoking an immune response. 

Particularly iDustrative polypeptides of the present invention comprise 
those aicoded by a polynucleotide sequence set forth in any one of SEQ ID NO;l-2234, 
or a sequence that hybridizes under moderately stringent conditions, or, alternatively, 

10 under highly stringent conditions, to a polynucleotide sequence set forth in any one of 
SEQ ID NO:l-2234. Certain other illustrative polypeptides of the invention comprise 
amino acid sequences as set forth in SEQ ID NO:223S. 

The polypeptides of the present invention are sometimes herein referred 
to as colon tumor proteins or colon tumor polypeptides, as an indication that their 

1 S identification has been based at least fai part upon then: increased levels of expression m 
colon tumor samples. Thus, a "colon tumor polypeptide** or "colon tumor protein," 
refers generally to a polypeptide sequence of the present invention, or a polynucleotide 
sequence encoding such a polypeptide, that is ^pressed in a substantial proportion of 
colon tumor samples, for example preferably greater than about 20%, more preferably 

20 greats than about 30%, and most preferably greater than about 50% or more of colon 
tumor samples tested, at a level that is at least two fold, and preferably at least five fold, 
greater than the level of expression in normal tissues, as determined using a 
representative assay provided herein. A colon tumor polypeptide sequence of the 
invention, based upon its increased level of expression in tumor cells, has particular 

25 utility both as a diagnostic marker as well as a therapeutic target, as further described 
below. 

In certain preferred embodiments, the polypeptides of the invention are 
immunogenic, they react detectably vdtbin an immunoassay (such as an ELISA or 
T-cell stimulation assay) with antisera and/or T-cells &om a patient with colon cancer. 
30 Screening for inraiunogenic activity can be performed usmg techniques well known to 
the skilled artisan. For example, such screens can be p^oimed using miethods such as 



BNSDOCID: <WO ^019638BAajL^ 



wo 01/96388 



9 



PCT/DS01/18S57 



those described in Harlow and Lane, Antibodies: A Laboratory Manual, Cold Spring 
Harbor Laboratory, 1988. In one illustradve exan:iple, a polypeptide may be 
immobilized on a solid support and contacted with patient sera to allow binding of 
antibodies within the sera to the hmnobilized polypeptide. Unbound sera may Aen be 

5 removed and bound antibodies detected using, for example, *^I-labeled Protdn A. 

As would be recognized by the skilled artisan, immunogenic portions of 
the polypeptides disclosed herein are also encompassed by the present invention. An 
''immunogenic portion," as used herein, is a fragment of an immunogenic polypeptide 
of the invention that itself is immunologically reactive (i.e., specifically binds) with the 

10 B-cells and/or T-cell surface antigen receptors that recognize the polypeptide. 
Immunogenic portions may generally be identified using well known techniques, such 
as those summarized in Paul, Fundamental Immunology, 3rd ed, 243*247 (Ravra Press, 
1993) and references cited therein. Such techniques include screening polypeptides for 
the ability to react with antigen-specific antibodies, antisera and/or T-cell lines or 

IS clones. As used herem, antisera and antibodies are "antigen-specific" if they 
^)ecifically bind to an antigen (i.e„ they react with the protein in an ELISA or other 
unmunoassay, and do not react detectably with unrelated proteins). Such antisera and 
antibodies may be prepared as described herein, and using well-known techniques. 

In one preferred embodiment, an inmiunogenic portion of a polypeptide 

20 of the present invention is a portion that reacts with antisera and/or T-cells at a level 
that is not substantially less than the reactivity of the fulHength polypeptide (e.g., in an 
ELISA and/or T-cell reactivity assay). Preferably, the level of immunogenic activity of 
flie immunogenic portion is at least about 50%, preferably at least about 70% and most 
preferably greater than about 90% of the immunogenicity for the fiill-Iength 

25 polypeptide. In some instances, preferred immunogenic portions will be identified that 
have a level of immunogenic activity greats than that of the corresponding fiill-length 
polypeptide, havmg greater than about 100% or 150% or more immunogenic 
activity. 

In certain other embodiments, illustrative immunogenic portions may 
30 include peptides in which an N-terminal leader sequence and/or transmembrane domain 
have been deleted. Other illustrative immunogenic portions wiU contain a small N- 



BNSOOCtCh <W0 P1963B8A9JL> 



wo 01/96388 



10 



PCTAJSOl/18557 



and/or C-terminal deletion {e.g., 1-30 amino acids, preferably 5-15 amino acids), 
relative to the mature protdn. 

In another embodiment, a polypeptide composition of the invention may 
also comprise one or more polypeptides that are immunologically reactive wth T cells 
5 and/or antibodies generated against a polypeptide of the invention, particularly a 
polypeptide having an amino acid sequence disclosed heran, or to an immunogenic 
fiagment or variant thereof. 

In another embodiment of the invention, polypeptides are provided that 
comprise one or more polypeptides fliat are enable of elidtrng T cells and/or 
10 antibodies fliat are immunologically reactive with one or more polypeptides described 
herein, or one or more polypeptides encoded by contiguous nucleic acid sequences 
contained in the polynucleotide sequences disclosed herein, of immunogenic fragments 
or variants fhereot or to one or more nucleic acid sequences which hybridize to one or 
more of these sequences under conditions of moderate to high stringency. 
1 5 The present invention, in another aspect, provides polypeptide fragments 

comprising at least about 5, 10, 15, 20, 25, 50, or 100 contiguous amino acids, or more, 
including all mtermediate lengths, of a polypeptide compositions set forth herein, such 
as those set forth m SEQ ID NO:2235, or those encoded by a polynucleotide sequence 
set forth in a sequence of SEQ ID NO:l-2234. 
20 In another aspect, tiie present invention provides variants of die 

polypeptide compositions described herein. Polypeptide variants generally 
encompassed by the present invention will typically exhibit at least about 70%, 75%, 
80%, 85%, 90%, 91%, 92%, 93%, 94%, 95%, 96%, 97%, 98%, or 99% or more identity 
(determined as described below), along its length, to a polypeptide sequences set forth 
25 herein* 

In one preferred embodiment, flie polypeptide fragments and variants 
provided by tiie present invention are immunologically reactive witii an antibody and/or 
T-cell that reacts with a full-length polypeptide specifically set forth herein. 

In another prefenred embodiment, the polypq)tide fragments and variants 
30 provided by the present invention exhibit a level of inmiunogenic activity of at least 
about 50%, preferably at least about 70%, and most preferably at least about 90% or 



BNSOOCID: <WO__0196888Aa_l_> 



wo 01/96388 



11 



PCTAJSOl/18557 



more of that exhibited by a full-length polypeptide sequence specifically set forth 
herein. 

A polypeptide Variant/' as the term is used herein, is a polypeptide that 
typically differs from a polypeptide specifically disclosed hosin in one or more 
5 substitutions, deletions, additions and/or insertions. Such variants may be naturally 
occurring or may be synthetically generated, for example, by modifymg one or more of 
the above polypeptide sequences of the invention and evaluating their immunogenic 
activity as described herein and/or using any of a number of techniques well known in 
the ait. 

10 For example, certain illustrative variants of the polypeptides of the 

mvention include those in which one or more portions, such as an N-terminal leader 
sequence or transmembrane domain, have been removed. Other illustrative variants 
include variants in which a small portion 1-30 amino acids, preferably 5-1 5 amino 
acids) has been removed firom the N- and/or C-terminal of the mature protein. 

15 In many instances, a variant will contain conservative substitutions. A 

"conservative substitution" is one in which an amino acid is substituted for another 
amino acid that has similar properties, such that one skilled in the art of peptide 
chemistry would expect the secondary structure and hydropathic nature of the 
polypeptide to be substantially unchanged. As described above, modifications may be 

20 made in the structure of the polynucleotides and polypeptides of the present invention 
and still obtain a ftinctional molecule that encodes a variant or derivative polypeptide 
with desirable characteristics, e.g., with immunogenic characteristics. When it is 
desired to alter the amino acid sequence of a polypeptide to create an equivalent, or 
even an improved, immunogenic variant or portion of a polypeptide of the invention, 

25 one skilled in the art will typically change one or more of the codons of the ^coding 
DN A sequence according to Table 1 . 

For example, c^tain amino acids may be substituted for other amino 
acids in a protein structure without q)preciable loss of interactive binding capacity with 
structures such as, for example, antigen-binding regions of antibodies or binding sites 

30 on substrate molecules. Since it is the interactive capacity and nature of a protein that 
defines that protein's biological functional activity, certain amino acid sequence 



(96388A3LL?> 



WQOl/96388 



PCT/USOl/18557 



12 

substitudons can be made in a protein sequence, and, of course, its underlying DNA 
coding sequence, and nevertheless obtain a protein with like properties. It is thus 
contemplated that various changes may be made in the peptide sequences of the 
disclosed compositions, or corresponding DNA sequences Avhich encode said peptides 
S without appiedable loss of their biological utility or activity. 

Table 1 



Amino Acids Codons 



Alanine 


Ala 


A 


GCA 


GCC 


GCG 


GCU 


Cysteine 


Cys 


C 


UGC 


UGU 






Aspaitic acid 


Asp 


D 


GAC 


GAU 






Ghitamic acid 


Glu 


E 


GAA 


GAG 






Phenylalanine 


Phe 


F 


UUC 


UUU 






Glycine 


Gly 


G 


GGA 


GGC 


GGG 


GGU 


Histidine 


His 


H 


CAC 


CAU 






Isoteucine 


lie 


I 


AUA 


AUC 


AUU 




Lysine 


Lys 


K 


AAA 


AAG 






Leucine 


Leu 


L 


XJUA 


UUG 


CUA 


cue 


Methionine 


Met 


M 


AUG 








Asparagine 


Asn 


N 


AAC 


AAU 






Proline 


Pro 


P 


CCA 


CCC 


CCG 


ecu 


Glutamine 


Gin 


Q 


CAA 


CAG 






Arginine 


Arg 


R 


AGA 


AGG 


CGA 


CGC 


Serine 


Ser 


S 


AGC 


AGU 


UCA 


ucc 


Threonine 


Tbr 


T 


ACA 


ACC 


ACG 


ACU 


Valnie 


Val 


V 


GUA 


GUC 


GUG 


GUU 


Tryptophan 


Trp 


W 


UGG 








Tyrosine 


Tyr 


Y 


UAC 


UAU 







CUG CUU 



CGG CGU 
UCG UCU 



10 In making such changes, the hydropathic index of amino acids may be 

considered. The importance of the hydropathic amino acid index in conferring 
interactive biologic function on a protein is generally understood in the art (Kyte and 



BNSDOCIft <WO ^0196388A2L1_> 



wo 01/96388 



13 



PCTAJSOl/18557 



Doolitfle, 1982, incoiporated herein by reference). It is accepted ftat the relative 
hydropathic character of the amino acid contributes to the secondary structure of the 
resultant protein, which in turn defines the interaction of the protein with oth^ 
molecules, for example, enzymes, substrates, receptors, DNA, antibodies, antigens, and 

5 the like. Each amino acid has been assigned a hydropathic index on the basis of its 
hydrophobicity and charge characteristics (Kyte and Doolittle, 1982). These values are: 
isoleucine (+4,5); valine (+4.2); leucine (+3.8); phenylalanine (+2.8); cysteine/cystine 
(+2.5); methionine (+1.9); alanme (+1.8); glycine (-^.4); threonine (-0.7); serine (~ 
0-8); tryptophan (-0.9); tyrosine (-1.3); proline (-1.6); histidine (-3.2); glutamate (- 

10 3.5); glutamine (-3.5); aspartate (-3.5); asparagine (-3.5); lysine (-3.9); and arginine (- 
4,5). 

It is known in flie art that certam amino acids may be substituted by 
other amino adds having a similar hydropathic index or score and still result in a 
protem with sunilar biological activity, le. still obtain a biolo^cal functionally 

IS equivalent piptein. In making such changes, the substitution of amino acids whose 
hydropathic indices are within 12 is inferred, those withm ±1 are particularly 
preferred, and tiiose wifein ±0.5 are even more particularly preferred. It is also 
understood in the art that the substitution of like amino adds can be made effectively on 
the basis of hydrophilicity. U. S. Patent 4,554,101 (specifically incorporated herein by 

20 reference in its entirety), states that the greatest local average hydrophilicity of a 
protein, as governed by the hydrophilicity of its adjacent amino acids, correlates with a 
biological property of the protein. 

As detailed in U. S. Patent 4,554,101, the following hydrophilicity 
values have been assigned to amino add residues: argmine (+3.0); lysine (+3.0); 

25 aspartate (+3.0 ± 1); glutamate (+3.0 ± 1); serine (+0.3); asparagine (+02); gjutamine 
(+0.2); glycine (0); Ihreonme (-0.4); proline (-0.5 ± 1); alanme (-0.5); histidme (-0.5); 
cysteme (-1.0); methionine (-13); valine (-1.5); leudne (-1.8); isoleucipe (-1-8); 
tyrosine (-2.3); phenylalamne (-2.5); tryptophan (-3.4). It is understood that an ammo 
acid can be substituted for another having a similar hydrophilicity value and still obtain 

30 a biologically equivalent, and m particular, an immunologically equivalent protein. In 
such changes, the substitution of amino acids whose hydrophilidty values are witinn i2 



BNSOOCID: <WO__0t96a8aAS!JU> 



WO01/9d388 



14 



PCTAJSOl/18557 



is preferred, those within ±1 arc particularly preferred, and those within ±0.5 are even 
more particularly preferred. 

As outlined above, amino acid substitutions are generally therefore based 
on the relative siniilarity of the amino add side-chain substituents, for example, their 

5 hydiophobicity, hydrophilicity, charge, size, and the like. Exemplary substitutions that 
take various of the foregoing characteristics into consideration are well known to those 
of skill in the art and include: argjnine and lysine; glutamate and aspartate; serine and 
threonine; glutamine and asparagine; and valine, leucme and isoleucine. 

In addition, any polynucleotide may be further modified to mcrease 

10 stability in vivo. Possible modifications include, but are not limited to, the addition of 
flanking sequences at the 5* and/or 3' ends; the use of phosphorothioate or T 0-methyl 
rather than phosphodiesterase Imkages m the backbone; and/or the mclusion of 
nontraditional bases such as inosine, queosine and wybutosine, as well as acetyl- 
methyl-, thio- and other modified forms of adenme, cytidine, guanine, Aymine and 

IS uridine. 

Amino add substitutions may further be made on the basis of similarity 
m polarity, charge, solubility, hydit)phobidty, hydrophilicity and/or the amphipalhic 
nature of the residues. For example, negativdy charged ammo acids include aspartic 
add and glutamic acid; positively charged amino acids mclude lysine and ai^ne; and 

20 amino acids with uncharged polar head groups having similar hydrophilicity values 
include leucine, isoleucine and valine; glycine and alanine; asparagme and glutanune; 
and serine, threomne, phenylalanine and tyrosine. Other groups of ammo acids that may 
represent conservative changes include: (1) ala, pro, gly, glu, asp, gin, asn, ser, thr; 
(2) cys, ser, tyr, thr, (3) val, ile, leu, met, ala, phe; (4) lys, arg, his; and (5) phe, tyr, tip, 

25 his. A variant may also, or alternatively, contain nonconservative changes. In a 
preferred embodiment, variant polypeptides differ from a native sequence by 
substitution, deletion or addition of five amino adds or fewer. Variants may also (or 
alternatively) be modified by, for example, the deletion or addition of amino acids that 
have mmimal mfluence on the unmunogemcity, secondary structure and hydropathic 

30 nature of the polypeptide. 



BNSDOdO: <WO___0186388A2_lj> 



wo 01/96388 



15 



PCTAJSOl/18557 



As noted above, polypeptides may comprise a signal (or leader) 
sequence at the N-terminal end of the protein, which co-transIationaQy or post- 
translationally directs transfer of tibie protein. The polypeptide may also be conjugated 
to a linker or other sequence for ease of synthesis, purification pr identijScation of the 

5 polypeptide (js,g., poly-His), or to enhance bindmg of the polypq)tide to a solid support. 
For example, a polypeptide may be conjugated to an immunoglobulin Fc region. 

When comparing polypeptide sequences, two sequences are said to be 
"identical" if the sequence of amino acids in the two sequences is the same when 
aligned for maximum correspondence, as described below. Comparisons between two 

10 sequences are typically performed by comparing the sequences over a comparison 
window to identify and compare local regions of sequence similarity. A "comparison 
window" as used herein, refers to a segment of at least about 20 contiguous positions, 
usually 30 to about 75, 40 to about SO, in which a sequence may be compared to a 
reference sequence of the same number of contiguous positions after the two sequences 

15 are optimally aligned. 

Optimal alignment of sequences for comparison may be conducted using 
the Megalign program in the Lasergene suite of bioinformatics software (DNASTAR, 
Inc., Madison, WI), using default parameters. This program embodies several 
alignment schones described in the following references: Dayhoff, M.O. (1978) A 

20 model of evolutionary change in proteins - Matrices for detecting distant relationships. 
In Dayhoff, M.O. (ed.) Atlas of Protein Sequence and Structure, National Biomedical 
Research Foundation, Washington DC Vol. 5, Suppl. 3, pp. 345-358; Hein J. (1990) 
Unified Approach to Alignment and Phylogenes pp. 626-645 Methods in Enzymology 
vol. 183, Academic Press, Inc., San Diego, CA; Higgins, D.G- and Sharp, P.M. (1989) 

25 CABIOS J:15M53; Myers, E.W. and Muller W. (1988) CABIOS 4:1 1-17; Robinson, 
E.D. (1971) Comb, Theor 11:105; Saitou, N. Nd, M. (1987) Mol Biol Evol 4:406- 
425; Sneath, P.H.A. and Sokal, RJR.. (1973) Numerical Taxonomy - the Principles and 
Practice of Numerical Taxonomy, Freeman Press, San Francisco, CA; Wilbur, W.J. and 
Lipman, DJ. (1983) Proc. Natl Acad, Sci. USA «0:726-730. 

30 Alternatively, optimal alignment of sequences for conq)arison may be 

conducted by the local identity algorithm of Smith and Watennan (1981) Add APL 



wo 01/96388 



16 



PCTAJSOl/18557 



' Math 2:482, by the identity alignment algorithm of Needleman and Wunsch (1970) J. 
MoL Biol 48:443, by the search for similarity methods of Pearson and Lipman (1988) 
Proc, Natl Acad, Set USA 85: 2444, by computerized implementations of these 
algorithms (GAP, BESTFIT, BLAST, FASTA, and TF ASIA m the Wisconsin Genetics 
5 Software Package, Genetics Computer Group (GCG), 575 Science Dr., Madison, WI), 
or by inspection. 

One preferred example of algorithms that are suitable for determining 
percent sequence identity and sequence similarity are the BLAST and BLAST 2.0 
algorisms, which are described m Altschul et al. (1977) NucL Acids Res. 25:3389-3402 

10 and Altschul et al. (1990) J. Mol Biol 215:403-410, respectively. BLAST and BLAST 
2.0 can be used, for example with the parameters described herein, to determine percent 
sequence identity for the polynucleotides and polypeptides of the invention. Software 
for performing BLAST analyses is publicly available through the National Center for 
Biotechnology Information. For amino acid sequences, a scoring matrix can be used to 

15 calculate the cumulative score. Extension of the word hits in each dhrection are halted 
when: the cumulative alignment score falls off by the quantiQr X from its maximum 
achieved value; the cumulative score goes to zero or below, due to the accumulation of 
one or more negative-scoring residue alignments; or the end of either sequence is 
reached. The BLAST algorithm parameters W, T and X determine the sensitivity and 

20 speed of the alignment. 

In one preferred approach, the '^percentage of sequence identity" is 
determined by comparing two optimally aligned sequences over a window of 
comparison of at least 20 positions, wh^in the portion of the polypeptide sequence in 
the comparison window may comprise additions or deletions (ie., gaps) of 20 percent 

25 or less, usually 5 to 15 percent, or 10 to 12 percent, as compared to the reference 
sequences (which does not comprise additions or deletions) for optimal aUgnment of the 
two sequences. The percentage is calculated by determining the nmnber of positions at 
which the identical amino acid residue occurs in both sequences to yield the number of 
matched positions, dividing the number of matched positions by the total number of 

30 positions in the reference sequence {ie., the window size) and multiplying fte results by 
100 to yield the percentage of sequence identity. 



BNSDOCID: ^WO__0196a8aA2LL> 



wo 01/96388 



17 



PCT/USOl/18557 



Within other ilhistrafive embodiments, a polypeptide may be a 
xCTOgehdc polypeptide flial comprises an polypeptide having substantial sequence 
identity, as described above, to Ate human polypeptide (also tenned autologous antigen) 
which served as a leference polypeptide, but which xenogeneic polypeptide is derived 

5 jErom a different, non-human spedcs. One skilled m the art will recognize that 
"self 'antigens are often poor stimulators of CD8+ and CD4+ T-lyraphocyte responses, 
and therefore efficient immunotherapeutic strategies directed against tumor 
polypeptides require the development of methods to overcome immune tolerance to 
particular self tumor polypeptides. For example, humans immunized with prostase 

10 protein fix)m a xenogeneic (non human) origm are capable of mounting an immune 
response against the counterpart human protein, e.g. the human prostase tumor protein 
present on human tumor cells. Accordmgly, the present invention provides methods for 
purifying the xenogeneic form of the tumor proteins set forth herein, such as the 
polypeptide set forfli in SEQ ID N02235, or those encoded by polynucleotide 

15 sequences set forth in SEQ ID NO:l-2234. 

Therefore, one aspect of flie present invention provides xenogeneic 
variants of the polypeptide compositiras described herein. Such xwiogeneic variants 
generally encompassed by the present invention will typically exhibit at least about 
70%, 75%, 80%, 85%, 90%, 91%, 92%, 93%, 94%, 95%, 96%, 97%, 98%, or 99% or 

20 moie identity along their lengflis, to a polypeptide sequences set forth herein. 

More particularly, the invention is directed to mouse, rat, monkey, 
porcine and other non-human polypeptides which can be used as xenogeneic forms of 
human, pplypeptides set forth herein, to mduce immune responses directed against 
tumor polypeptides of the invention. 

25 Within other illustrative embodiments, a polypeptide may be a fosion' 

polypeptide that ccHnprises multiple polypeptides as described herein, or that comprises 
at least one polypeptide as described herdn and an unrelated sequence, such as a known 
tumor protdn, A fusion partner may, for example, assist in providing T helper epitopes 
(an hrimunological fiiaon partner), prefembly T helper epitopes^ recognized by humans, 

30 or may assist m expressing the protein (an expression enhancer) at higher yields than 
the native recombinant protein. Certain preferred fusion partners are both 



BNSDOCIO: <WO__J[HW388A2_l> 



wo 01/96388 



18 



PCT/USOl/18557 



uiununoiogical and expression enhancing fusion partners. Other fusion partners may be 
selected so as to increase the solubili^ of the polypeptide or to enable the polypeptide 
to be targeted to desired intracellular compartments. Still further fusion partners 
include afiBnity tags» which fadlitate purification of the polypeptide. 
S Fusion polypeptides may generally be prepared using standard 

techniques, including chemical conjugation. Preferably, a fiision polypeptide is 
expressed as a recombinant polypeptide, allowing the production of increased levels, 
relative to a non-fused polypeptide, in an expression system. Briefly, DNA sequences 
encoding the polypeptide components may be assembled separately, and ligated into an 
10 appropriate ^pression vector. The 3' end of the DNA sequence encoding one 
polypeptide component is ligated, with or without a peptide linker, to the 5' end of a 
DNA sequence encoding the second polypeptide component so that the reading frames 
of the sequences are in phase. This permits translation into a single fusion polypeptide 
that retains the biological activity of both componrat polypeptides. 
IS A peptide linker sequence may be employed to separate the first and 

second polypeptide components by a distance sufficient to ensure that each polypeptide 
folds into its secondary and tertiary structures. Such a peptide linker sequence is 
incorpojBted into the fusion polypeptide using standard techniques well known in the 
art Suitable peptide linker sequences may be chosen based on the following factors: 
20 (1) their ability to adopt a flexible extended conformation; (2) their inability to adopt a 
secondary structure that could interact with functional epitopes on the first and second 
polypeptides; and (3) the lack of hydrophobic or charged residues that might react with 
the polypeptide functional epitopes. Preferred peptide linker sequences contain Gly, 
Asn and Ser residues. Other near neutral amino acids, such as Thr and Ala may also be 
25 used in the linker sequence. Amino acid sequences which may be usefully employed as 
linkers include those disclosed in Maratea et al.. Gene 40:39-46, 1985; Murphy et al., 
Proc Natl. Acad ScL USA 53:8258-8262, 1986; U.S. Patent No. 4,935,233 and U.S. 
Patent No. 4,751,180. The linker sequence may generally be from 1 to about 50 amino 
acids in length. Linker sequences are not required when the first and second 
30 polypeptides have non-essential N-tenninal amino acid regions that can be used to 
separate the functional domauis and prevent steric intaference. 



BNSDCX»0: <W0 P186388A?_L> 



wo 01/96388 



19 



PCT/USOl/18557 



The ligated DNA sequences are operably linked to suitable 
transcriptional or translational regulatory elements. The regulatory elements 
responsible for e^qiression of DNA are located only 5* to the DNA sequence encoding 
the first polypeptides. Similarly, stop codons required to end translation and 

5 transcription termination signals are only present 3* to the DNA sequaice encoding the 
second polypeptide. 

The fusion polypeptide can comprise a polypeptide as described herein 
together with an unrelated immunogenic protein, such as an immunogenic protein 
capable of eliciting a recall response. Examples of such proteins include tetanus, 

10 tuberculosis and hepatitis proteins {see^ for example, Stoute etal. New Engl J. Med.^ 
53(5:86-91, 1997). 

In one preferred embodiment, the immunological fusion partner is 
derived firom a Mycobacterium sp., such as a Mycobacterium tuberculosis-derived Ral2 
fragment. Ral2 compositions and methods for thdur use hfi enhancing the expression 

15 and/or immunogenicity of heterologous polynucleotide/polypeptide sequences is 
described in U.S. Patent Application 60/158,585, the disclosure of which is 
incoiporated hercm by reference in its entirety. Briefly, Ral2 refers to a polynucleotide 
region that is a subsequence of n Mycobacterium tuberculosis MTB32A nucleic acid. 
MTB32 A is a serine protease of 32 KD molecular wdgiht encoded by a gene in virulent 

20 and avirulent strains of M. tuberculosis. The nucleotide sequence and amino acid 
sequence of MTB32A have been described (for example, U.S. Patent Application 
60/158,585; see also, Skeiky et a/., Infection and Immun, (1999) 67:3998-4007, 
incorpomted herein by reference). C-terminal fragments of the MTB32A coding 
sequence express at high levels and remain as a soluble polypeptides throughout the 

25 purification process. Moreover, Ral2 may enhance the immunogemcity of heterologous 
hnmunogenic polypeptides with v^hich it is fused. One prefrared Ral2 fusion 
polypeptide comprises a 14 KD C-terminal fragment corresponding to amino acid 
residues 192 to 323 of MTB32A. Other preferred Ral2 polynucleotides genially 
comprise at least about 15 consecutive nucleotides, at least about 30 nucleotides, at 

30 least about 60 nucleotides, at least about 100 nucleotides, at least about 200 nucleotides, 
or at least about 300 nucleotides that encode a portion of a Ral2 polypeptide. Ral2 



BNSDOCtO: <VWO ^019638aA^L>' 



PCTAJS01/18SS7 

20 

polynucleotides may comprise a native sequence {Le., an endogenous sequence that 
encodes a Ral2 polypeptide or a portion thereof) or may comprise a variant of such a 
sequence. Ral2 polynucleotide variants may contain one or more substitutions^ 
additions, deletions and/or insertions such that the biological activity of tiie mcoded 
S fusion polypeptide is not substantially diminished, relative to a fusion polypeptide 
comprising a native Ral2 po]ypq)tide. Variants preferably exhiWt at least about 70% 
identity, more preferably at least about 80% identity and most preferably at least about 
90% identity to a polynucleotide sequence that encodes a native Ral2 polypeptide or a 
portion thereof. 

10 Within other preferred embodiments, an immunological fusion partner is 

derived jfrom protem D, a surface protein of the gram-negative bacterium Haemophilus 
influenza B (WO 91/18926). Preferably, a protein D derivative comprises 
approximately the first third of the protem (e.g„ the first N-terminal 100-110 amino 
adds), and a protdn D derivative may be lipidateA Within certain preferred 

15 embodiments, the first 109 residues of a Lipoprotein D fusion partner is mcluded on the 
N-terminus to provide the polypeptide with additional exogenous T-cell epitopes and to 
increase the expression level in K coli (thus fimctioning as an expres^on enhancer). 
The lipid tail ensures qptunal presentation of the antigen to antigen presentmg cells. 
Other fusion partners mclude the non-structrarf protein from influenzae virus, NSl 

20 Oiemaglutinin). Typically, the N-termmal 81 amuio acids axe used, although different 
fi-agments that include T-helper epitopes may be used. 

In another embodiment, the immunological fusion partner is the protein 
known as LYTA, or a portion thereof (preferably a C-terminal portion). LYTA is 
derived from Streptococcus pneumoniaey which synthesizes an N-acetyl-L-alanine 

25 amidase known as amidase LYTA (encoded by the LytA gene; Gene 45:265-292, 
1986). LYTA is an autolysin that specifically degrades certain bonds in the 
peptidoglycan backbone. The C-terminal domain of the LYTA protein is responsible 
for the afOnity to the choline or to some choline analogues such as DEAE. This 
property has been exploited for the development of E. coli C-LYTA expressing 

30 plasmids useful for expression of fusion proteins. Purification of hybrid proteins 
containing the C-LYTA fragment at the amino t^minus has been described (see 



WO 01/96388 



BNSOOCIO: *W0 ^0l96a8aA2LL> 



wo 01/96388 



21 



PCT/USOl/18557 



Biotechnology 70:795-798, 1992). Within a preferred embodiment, a repeat portion of 
LYTA may be incorporated into a fusion polypeptide. A repeat portion is found in the 
C-tenninal region starting at residue 178. A particularly preferred repeat portion 
incorporates residues 1 88-305. 

5 Yet another illustrative embodiment involves fusion polypeptides, and 

tiie polynucleotides encoding them, wherein the fusion partner comprises a targeting 
signal capable of directing a polypeptide to the endosomaWysosomal con^artment, as 
described in U.S. Patent No. 5,633,234. An immunogenic polypeptide of the invention, 
when fused with this targeting signal, will associate more efficiently with MHC class 11 

10 molecules and thereby provide enhanced in vivo stimulation of CDA^ T-cells specific 
for the polypeptide. 

Polypeptides of the invention are prepared using any of a variety of well 
known synflietic and/or recombinant techniques, the latter of which are further 
described below. Polypeptides, portions and other variants generally less than about 

15 150 amino adds can be generated by synthetic means, usii^ techniques well known to 
those of ordinary skill in the art In one illustrative example, such polypq)tides are 
synthesized using any of the conmiercially available solid-phase techniques, sudi as the 
Merrifield solid-phase synthesis method, where amino acids are sequentially added to a 
growing amino acid chain. See Menifield, J. Am ChenL Soc «J:2149-2146, 1963. 

20 Equipment for automated synthesis of polypeptides is commercially available from 
suppliers such as Perkin Elmer/Applied BioSystems Division (Foster City, CA), ard 
may be operated according to the manufacturer's instructions. 

In general, polypeptide compositions (including fusion polypeptides) of 
the invention are isolated. An "isolated'' polypeptide is one that is removed jfrom its 

25 original environment. For example, a naturally-occurring protein or polypeptide is 
isolated if it is separated fiom some or all of the coexisting materials in the natural 
system. Preferably, such polypq)tides are also purified, e.g., are at least about 90% 
pure, more preferably at least about 95% pure and most preferably at least about 99% 
pure. 



BNSDOaO: <WO P1B6388A9JL^ 



wo 01/96388 



22 



PCT/US01/18SS7 



Polynucleotide Compositions 

The present invention, in other aspects, provides polynucleotide 
compositions. The terms "DNA" and "polynucleotide" are used essentially 
interchangeably herein to refer to a DNA molecule that has been isolated free of total 

S genomic DNA of a particular species. "Isolated," as used herein, means that a 
polynucleotide is substantially away from other coding sequences, and that the DNA 
molecule does not contain large portions of unrelated codmg DNA, such as large 
chromosomal fragments or other frmctional graes or polypeptide coding regions. Of 
course, this refers to the DNA molecule as originally isolated, and does not exclude 

1 0 genes or coding regions later added to the segment by the hand of man. 

As will be understood by those skilled in the art, the polynucleotide 
compositions of this invention can include genomic sequences, extra-genomic and 
plasmid-encoded sequences and smaller engineered gene segments that express, or may 
be adapted to express, proteins, polypeptides, peptides and the like. Such segments 

1 5 may be naturally isolated, or modified synthetically by the hand of man. 

As will be also recognized by the skilled artisan, polynucleotides of the 
invention may be smgle-stranded (coding or antisense) or double-stranded, and may be 
DNA (genomic, cDNA or synthetic) or BNA molecules. RNA molecules may include 
HnRNA molecules, which contain mtrons and correspond to a DNA molecule in a one- 

20 to-one manner, and mRNA molecules, which do not contain introns. Additional coding 
or non-coding sequences may, but need not, be present within a polynucleotide of the 
present invention, and a polynucleotide may, but need not, be linked to other molecules 
and/or support materials. 

Polynucleotides may comprise a native sequence (/.e., an endogenous 

25 sequence that encodes a polypeptide/protein of the invention or a portion thereof) or 
may comprise a sequence that encodes a variant or derivative, preferably and 
immunogenic variant or derivative, of such a sequence. 

Therefore, accordmg to another aspect of the present invention, 
polynucleotide compositions are provided that comprise some or all of a polynucleotide 

30 sequence set forth in any one of SEQ ID NO:l*2234, complements of a polynucleotide 
sequence set forth m any one of SEQ ID NO: 1-2234, and degenerate variants of a 



BNSDOCID: <WO___019e388AiJL> 



wo 01/96388 



23 



PCTAJSOl/18557 



polynucleotide sequence set forth in any one of SEQ ID NO:l-2234. In certain 
preferred embodiments, the polynucleotide sequences set forth herein encode 
immunogauc polypeptides, as described above. 

In other related embodiments, die present^Jnvention provides 

5 polynucleotide variants havmg substantial identity to die sequences disclosed herem in 
SEQ ID NO: 1-2234, for example ttiose comprismg at least 70% sequence identity, 
preferably at least 75%, 80%, 85%, 90%, 95%, 96%, 97%, 98%, or 99% or higher, 
sequence identity compared to a polynucleotide sequence of this invention using the 
methods described herein, {e,g,y BLAST analysis using standard parameters, as 

10 described below). One skilled in this art will recognize that these values can be 
appropriately adjusted to determine corresponding identity of protems encoded by two 
nucleotide sequences by taking into account codon degeneracy, ammo acid similarity, 
reading frame positioning and fhe like. 

Typically, polynucleotide variants will contain one or more substitutions, 

15 additions, deletions and/or insertions, preferably such Aat the immunogenicity of the 
polypeptide encoded by the variant polynucleotide is not substantially diminished 
relative to a polypeptide encoded by a polynucleotide sequence specifically set forth 
herem). The term 'Variants" should also be understood to encompasses homologous 
genes of xenogeneic origin. 

20 In additional embodiments, the present invention provides 

polynucleotide fragments comprising or consisting of various lengths of contiguous 
stretches of sequence identical to or complementary to one or more of the sequences 
disclosed ^herein. For example, polynucleotides are provided by this invention that 
comprise or consist of at least about 10, 15, 20, 30, 40, 50, 75, 100, 150, 200, 300, 400, 

25 500 or 1000 or more contiguous nucleotides of one or more of the sequences disclosed 
herein as well as all intermediate lehgdis there between. It will be readily understood 
that "mtermediate lengths", in this context, means any lengfli between the quoted 
values, such as 16, 17, 18, 19, etc.; 21, 22. 23, etc.; 30, 31, 32, etc.; 50, 51, 52, 53, etc.; 
100, 101, 102, 103, etc.; 150, 151, 152, 153, etc.; including all integers through 200- 

30 500; 500*1,000, and the like. A polynucleotide sequence as described here may be 
extended at one or both ends by additional nucleotides not found in the native sequence. 



BNSDOCID:<WO. 



196388A2JL^ 



WOOl/96388 PCT/USOl/18557 

24 

This additional sequence may consist of 1, 2, 3, 4, 5, 6, 7, 8, 9, 10, 11, 12, 13, 14, 15, 
16, 17, 18, 19, or 20 nucleotides at eiSaesr end of the disclosed sequence or at both ends 
of the disclosed sequence. 

In another embodiment of the invention, polynucleotide compositions 

5 are provided that are capable of hybridizing under moderate to high stringency 
conditions to a polynucleotide sequence provided ho^in, or a fragment thereof, or a 
complementary sequence thereof Hybridization techniques are well known in the art 
of molecular biology. For purposes of illustration, suitable moderately stringent 
conditions for testing the hybridization of a polynucleotide of this invention with other 

10 polynucleotides include prewashing in a solution of 5 X SSC, 0.5% SDS, 1.0 mM 
EDTA (pH 8.0); hybridizing at 50°C-60^C, 5 X SSC, overnight; followed by washing 
twice at Sy'C for 20 minutes with each of 2X, 0,5X and 0.2X SSC containing 0.1% 
SDS* One skOIed in the art will understand that the stringency of hybridization can be 
readily manipulated, such as by altering the salt content of tfie hybridization solution 

15 and/or fte temperature at which the hybridization is performed. For example, in 
another embodiment, suitable higihly stringent hybridization conditions include those 
described above, with the exception that the temperature of hybridization is increased, 
e.g., to eO^S^'C or 65-7Q''C. 

In certain preferred embodiments, the polynucleotides described above, 

20 e.g., polynucleotide variants, fragments and hybridizing sequences, encode 
polypeptides that are immunologically cross-reactive with a polypeptide sequence 
specifically set forth herein. In other preferred embodiments, such polynucleotides 
encode polypeptides that have a level of immunogenic activity of at least about 50%, 
preferably at least about 70%, and more preferably at least about 90% of that for a 

25 polypeptide sequence specifically set forth her^ 

The polynucleotides of the presmt invention, or fragments thereof, 
regardless of the length of the coding sequence itself, may be combined with other 
DNA sequences, such as promoters, polyadenylation signals, additional restriction 
enzyme sites, multiple cloning sites, ptiier coding segments, and the like, such that their 

30 overall length may vary considerably. It is therefore contemplated that a nucleic add 
firagment of almost any lengdi may be employed, with the total length preferably being 



wo 01/96388 



25 



PCTAJSOl/18557 



limited by the ease of preparation and use in the intended recombmant DNA protocol. 
For example, illustrative polynucleotide segments with total lengths of about 10,000, 
about 5000, about 3000, about 2,000, about 1,000, about 500, about 200, about 100, 
. about 50 base pairs in length, and the like, (including all mtemediate lengths) are 
5 contemplated to be useful in many implementations of this inveirtion. 

When comparing polynucleotide sequences, two sequences are said to be 
"identical'* if the sequence of nucleotides in the two sequences is the same when aligned 
for maximum correspondence, as described below. Comparisons between two 
sequences are typically performed by comparing the sequences ova: a comparison 
10 window to identify and compare local regions of sequence similarity. A "coniparison 
window" as used herein, refers to a segment of at least about 20 contiguous positions, 
usually 30 to about 75, 40 to about 50, in which a sequence may be compared to a 
reference sequence of the same number of contiguous positions after the two sequences 
are optnnally aligned. 

1 5 Optimal alignment of sequences for comparison may be conducted using 

the Megalign program in the Lasergene suite of biomformatics software (DNASTAR, 
be, Madison, WI), using default parameters. This program embodies several 
alignment schemes described m the foDowing references: Dayhoflf, M.O. (1978) A 
model of evolutionary change in proteins - Matrices for detecting distant relationships. 

20 In Dayhoff, M.O. (ed.) Atlas of Protem Sequence and Structure, National Biomedical 
Research Foundation, Washington DC Vol. 5, Suppl. 3, pp. 345-358; Hdn J. (1990) 
Unified Approach to Alignment and Phylogenes pp. 626-645 Methods in Emymology 
vol. 183, Academic Press, hic, San Diego, CA; Higgins, D.G. and Sharp, PJs4. (1989) 
CABIOS 5:151-153; Myers, E.W. and ikvSkx W. (1988) CABIOS 4:11-17; Robmson, 

25 E.D. (1971) Comb. Theor 7i:105; Santou, N. Nes, M. (1987) Mol Biol Evol 4:406- 
425; Sneath, ?JiA. and Sokal, R.R. (1973) Numerical Taxonomy ^ the Principles and 
Practice of Numerical Taxonomy, Freeman Press, San Francisco, CA; Wilbur, W.J. and 
Lipman, DJ. (1993) Proc. Natl Acad, Scl USA W:726-730. 

Alternatively, optimal alignment of sequences for comparison may be 

30 conducted by the local identity algorithm of Smift and Waterman (1981) Add APL 
Math 2:482, by the identity alignment algorithm of Needleman and Wunsch (1970) J,- 



BNSDOCID: <WO ^Dige38aMLL> 



wo 01/96388 



26 



PCTAJS01/18S57 



MoL Bid 48:443, by the search for shnilarity methods of Pearson and Lipman (1988) 
Proc, Natl. Acad ScL USA 85: 2444, by computerized implementations of these 
algorithms (GAP, BESTFTT, BLAST, FASTA, and TFASTA in the Wisconsin Genetics 
Software Package, Genetics Compute Group (GCG), 575 Science Dr., Madison, WI), 
S or by inspection. 

One preferred example of algorithms that are suitable for determining 
percent sequence identity and sequence shnilarity are the BLAST and BLAST 2.0 
' algorithms, which are described in Altschul et al. (1977) NucL Acids Res. 25:3389-3402 
and Altschul et al. (1990) J. MoL Biol 215:403-410, respectively. BLAST and BLAST 

10 2.0 can be used, for example with the parameters described herein, to detennme percent 
sequence identity for the polynucleotides of the invention. Software for performing 
BLAST analyses is publicly available through the National Center for Biotechnology 
Information. In one illustrative example, cumulative scores can be calculated usmg, for 
nucleotide sequences, the parameters M (reward score for a pair of matching residues; 

15 always >0) and N (penalty score for mismatching residues; always <0). Extension of 
the word hits in each direction are halted when: the cumulative alignment score falls off 
by tiie quantity X from its maximum achieved value; the cumulative score goes to zero 
or below, due to the accumulation of one or more negative-scoring residue alignments; 
or the end of either sequence is reached; The BLAST algorithm parameters W, T and X 

20 deteimine the s^isitivity and speed of the alignment. The BLASTN program (for 
nucleotide sequences) uses as defaults a wordlength (W) of 11, and expectation (E) of 
10, and the BLOSUM62 scoring matrix (see Henikoff and Henikoff (1989) Proc, Natl 
Acad Set USA 89:10915) aUgranents, (B) of 50, expectation (E) of 10, M=5, N=-4 and 
a comparison of both strands. 

25 Preferably, the '*perc^tage of sequence identity" is determined by 

comparing two optimally aligned sequences over a window of comparison of at least 20 
positions, wh^in die portion of the polynucleotide sequence in ttie comparison 
wmdow may comprise additions or deletions {ie., gaps) of 20 percent or less, usually 5 
to 15 percent, or 10 to 12 percent, as compared to the reference sequences (which does 

30 not comprise additions or deletions) for optimal alignment of the two sequences. The 
percentage is calculated by determining the number of positions at which the identical 



BNSDOCIO: <W0 P19638BA?.L> 



wo 01/96388 



27 



PCT/USOl/18557 



nucldc acid bases occurs in both sequences to yield the number of matched positions, 
dividing the number of matched positions by the total number of positions in the 
reference sequence (ie., the window size) and multiplying the results by 100 to yield 
tiie percentage of sequence identity. 

S It will be appreciated by those of ordinary skill in the art Ifaat, as a result 

of the degeneracy of the genetic code, there are many nucleotide sequences that encode 
a polypeptide as described herein. Some of fliese polynucleotides bear minimal 
homology to the nucleotide sequence of any native gene. Nonetheless, polynucleotides 
that vary due to differences in codon usage are specifically contemplated by the present 

10 invention. Further, alleles of the genes comprising the polynucleotide sequences 
I»ovided herem are within the scope of the preset invention. Alleles are endogenous 
genes that are altered as a result of one or more mutations, such as deletions, additions 
and/or substitutions of nucleotides. The resulting mRNA and protein may, but need 
not, have an altered structure or function. Alleles may be identified using standard 

1 S techniques (such as hybridization, amplification and/or database sequence conapaiison). 

Therefore, in another embodiment of the invention, a mutagenesis 
approadi, such as site-specific mutagenesis, is employed for the preparation of 
immunogenic variants and/or derivatives of the polypeptides described herein* By this 
approach, specific modifications in a polypeptide sequence can be made through 

20 mutagenesis of the underlying polynucleotides that encode them. These techniques 
provides a straightforward approach to prepare and test sequence variants, for example, 
incorporating one or more of the foregoing considerations, by introducing one or more 
nucleotide sequence changes into the polynucleotide. 

Site-specific mutagenesis allows the i»roduction of mutants through the 

2S use of specific oligonucleotide sequences which encode tiie DNA sequence of die 
desired mutation, as well as a sufScient number of adjacent nucleotides, to provide a 
primer sequence of sufficient size and sequence complexity to form a stable duplex on 
both sides of the deletion junction being traversed. Mutations may be employed in a 
selected polynucleotide sequence to improve, alter, decrease, modify, or otiierwise 

30 change the properties of die polynucleotide itself, and/or alter flie properties, activity, 
compoidtion, stability, or primary sequence of the encoded polypeptide. 



BNSDOCIO: 



l9638aAaj^ 



wo 01/96388 



28 



PCTAJSOl/18557 



In certain embodiments of the present invention, the inventors 
contemplate the mutagenesis of the disclosed polynucleotide sequences to alter one or 
more properties of ihe encoded polypeptide, such as the inununogenidty of a 
polypeptide vaccine. The techniques of site-specific mutagenesisare ivell-known in the 

S art, and are widely used to create variants of boft polypeptides and polynucleotides. 
For example, site-specific mutagenesis is often used to altar a specific portion of a DNA 
molecule. In such embodiments, a primer comprising typically about 14 to about 25 
nucleotides or so in length is employed, with about 5 to about 10 residues on both sides 
of the junction of the sequence being altered. 

10 As will be appreciated by those of skill in the art, site-specific 

mutagenesis techniques have often employed a phage vector that exists in both a single 
stranded and double stranded form. Typical vectors useful in site-directed mutagenesis 
include vectors sudi as the M13 phage. These phage are readily 
commercially-available and their use is generally well-known to those skilled in the art 

IS Double-stranded plasmids are also routinely employed in site directed mutagenesis that 
eliminates the step of transferring the gene of interest fiom a plasmid to a phage. 

In general, dte-directed mutagenesis in accordance herewitti is 
poformed by first obtaining a single-stranded vector or melting apart of two strands of 
a double-stranded vector that includes within its sequence a DNA sequence that 

20 encodes the desired peptide. An oligonucleotide primer bearing the desired mutated 
sequence is prepared, generally synthetically. This primer is then annealed with the 
single-stranded vector, and subjected to DNA polymerizing enzymes such as E. coli 
polymerase I K16now fragment, in order to complete the synthesis of the mutation- 
bearing strand. Thus, a heteroduplex is formed wherein one strand encodes the original 

25 non-mutated sequence and the second strand bears the desired mutation. This 
heterodt9)lex vector is then used to transform appropriate cells, sudi as £ coli cells, and 
clones are selected which include recombinant vectors bearing the mutated sequence 
arrangement. 

The preparation of sequence variants of the selected peptide-encoding 
30 DNA segmrats using site-directed mutagenesis provides a means of producing 
potentially useful species and is not meant to be limiting as there are other ways in 



BNSDOCID: <W0 ^0196388A2JL> 



wo 01/96388 



29 



PCt/USOl/18557 



which sequence variants of peptides and flie DNA sequOTces encoding them may be 
obtained. For example, recombinant vectors encoding fte desired peptide sequence 
may be treated with mutagenic agents, such as hydioxylamine, to obtain sequence 
variants. Specific details regarding these methods and protocols are found m the 
5 teachings of Maloy etal, 1994; Segal, 1976; Prokop and Bajpai, 1991; Kuby, 1994; 
and Maniatis et al, 1982, each incorporated herein by reference, for that purpose. 

As used herein, the term "oligonucleotide directed mutagenesis 
procedure" refers to template-dependent processes and vector-mediated propagation 
which result in an increase m the concentration of a specific nucleic acid molecule 
10 relative to its initial concentration, or in an increase in the concentration of a detectable 
signal, such as amplification. As used herein, the tenn "oligonucleotide directed 
mutagenesis procedure" is intended to refer to a process that involves the 
template-depmdent extension of a primer molecule. The term template dependent 
process refers to nucleic acid synthesis of an RNA or a DNA molecule wherein the 
15 sequence of the newly synthesized strand of nucleic add is dictated by the well-known 
rules of complementary base pairing (see, for example, Watson, 1987). Typically, 
vector mediated methodologies involve the introduction of the nucleic acid fiagment 
into a DNA or RNA vector, the clonal amplification of the vector, and the recovery of 
the amplified nucleic acid fragment. Examples of such methodologies are provided by 
20 U. S. Patent No. 4,237,224, specifically incorporated herein by reference in its entirety. 

In another approach for the production of polypeptide variants of the 
present mvention, recursive sequence recombination, as described in U.S. Patent No. 
5,837,458^ may be employed. In this approadi, iterative cycles of recombination and 
screening or selection are performed to "evolve^* mdividual polynucleotide variants of 
25 the invention having for example, enhanced unmunogenic activity. 

In other embodiments of the present invention, the polynucleotide 
sequences provided herein can be advantageously used as probes or primers for nucleic 
acid hybridization. As such, it is contemplated that nucleic add segments that comprise 
or consist of a sequence region of at least about a 15 nucleotide long contiguous 
30 sequence that has the same sequence as, or is complementary to, a 15 nucleotide long 
contiguous sequence disclosed herein will find particular utility. Longer contiguous 



BNSOOaO: <WO___01fl6388A«_L> 



wo 01/96388 



30 



PCTAJSOl/18557 



identical or complementary sequences, e.g., those of about 20, 30, 40, 50, 100, 200, 
500, 1000 (including all intemediate lengths) and even up to full length sequences will 
also be of use in certain embodiments. 

The ability of such nucleic acid probes to specifically hybridize to a 

5 sequence of interest wUI enable them to be of use in detecting the presence of 
complementary sequences in a given sample. However, other uses are also envisioned, 
such as the use of the sequence information for the prepaiation of mutant species 
primers, or primers for use in preparing other genetic constructions. 

Polynucleotide molecules having sequence regions consisting of 

10 contiguous nucleotide stretches of 10-14, 15-20, 30, 50, or even of 100-200 nucleotides 
or so (including intermediate lengths as well), identical or complementary to a 
polynucleotide sequence disclosed herein, are particularly contemplated as 
hybridization probes for use in, e.g.. Southern and Northern blotting. This would allow 
a gene product, or fragment thereof, to be analyzed, both in diverse cell types and also 

IS in various bacterial, cells. Hie total size of fragment, as well as the size of tiiie 
complementary slretch(es), will ultimately depend on ibs mtended use or application of 
the particular nucleic acid segment. Smaller fragments will generally find use in 
hybridization embodiments, wherem &e length of the contiguous complementary 
region vasy be varied, such as between about 15 and about 100 nucleotides, but larger 

20 contiguous complementarity stretches may be used, according to the length 
complementary sequences one wishes to detect 

The use of a hybridization probe of about 15-25 nucleotides in length 
allows the formation of a duplex molecule that is both stable and selective. Molecules 
having contiguous complementary sequences over stretches greater than 15 bases in 

25 length are generally preferred, though, in order to increase stability and selectivity of 
the hybrid, and thereby improve the quality and degree of specific hybrid molecules 
obtained. One will generally prefer to design nucleic acid molecules having gene- 
complementary stretches of 15 to 25 contiguous nucleotides, or even longer where 
desired. 

30 Hybridization probes may be selected from any portion of any of the 

sequences disclosed herein. All that is reqmred is to review the sequences set forth 



BHSDOOD; ^WO pig6388A^_L> 



wo 01/96388 PCTAJSOl/18557 

31 

herein, or to any continuous portion of the sequences, from about 15-25 nucleotides in 
length up to and mcluding the fiill length sequence, that one vdshes to utilize as a probe 
or primer. The choice of probe and primer sequences may be govOTied by various 
fectors. For example, one may wish to employ primers from towards the termini of tiie 
5 total sequence. 

Small polynucleotide segments or fragments may be readily prepared by, 
for example, dkectly synthesizing the fragment by chemical means, as is commonly 
practiced using an automated oligonucleotide synthesizer. Also, fragments may be 
obtained by application of nucleic acid reproduction technology, such as the PGR™ 
10 technology of U. S. Patent 4,683,202 (incorporated herein by reference), by introducing 
selected sequences into recombinant vectors for recombinant production, and by other 
recombinant DNA techniques generally known to those of skill in the art of molecular 
biology. 

The nucleotide sequences of the invention may be used for their ability 

15 to selectively form duplex molecules with complementary stretohes of tiie entire gene or 
gene fragments of interest Depending on the application envisioned, one will typically 
desire to employ varying conditions of hybridization to achieve varying degrees of 
selectivity of probe towards target sequence. For applications requiring high 
selectivity, one will typically desire to employ relatively stringent conditions to form 

20 the hybrids, e.g., one will select relatively low salt and/or high temperature conditions, 
4 such as provided by a salt concentration of from about 0.02 M to about 0.15 M salt at 
temperatures of from about SO'^C to about TO^'C. Such selective conditions tolerate 
little, if apy, mismatch between the probe and the template or target strand, and would 
be particularly suitable for isolating related sequences. 

25 Of course, for some applications, for exaniple, where one desires to 

prepare mutants employing a mutant primer strand hybridized to an underiying 
template, less stringent (reduced stringency) hybridization conditions will typically be 
needed in order to allow formation of the heterodiq>Ira. In these circumstances, one 
may desire to employ salt conditions such as those of from about 0.15 M to about 0.9 M 

30 salt, at temperatures ranging from about 20''C to about 55^C. Cross-hybridizing species 
can thereby be readily idratified as positively hybridizing signals with respect to control 



BNSDOCIO: <WO ^019838aA2^L> 



wo 01/96388 



32 



PCT/USOl/18557 



hybridizations. In any case, it is generally appreciated that conditions can be rendered 
more stringent by the addition of increasing amounts of formamide, >vhich serves to 
destabilize the hybrid duplex in the same manner as increased temperature. Thus, 
hybridization conditions can be readily manipulated, and thus will generally be a 

S method of choice depmding on the desired results. 

According to another embodiment of the present invention, 
polynucleotide compositions comprising antisense oligonucleotides are provided. 
Antisense oligonucleotides have been demonstrated to be effective and targeted 
inhibitors of protein syntibesis, and, consequently, provide a therapeutic approach by 

10 which a disease can be treated by inhibiting the synthesis of proteins that contribute to 
the disease. The efficacy of antisense oligonucleotides for inhibiting protein synthesis 
is well established. For example, the synthesis of polygalactaiironase and the muscarine 
type 2 acetylcholine receptor are inhibited by antisense oligonucleotides directed to 
their respective mRNA sequences (U. S. Patent 5,739,1 19 and U. S. Patent 5,759,829). 

15 Further, examples of antisense inhibitton have been demonstrated with the nuclear 
protdn cyclin, the multiple drug resistance gme (MDGl), ICAM-1, E-selectin, STK-1, 
striatal GABAa receptor and human EOF (Jaskulski et a/.. Science. 1988 Jun 
10;240(4858):1544-6; Vasanthakumar and Ahmed, Cancer Commun. 1989;I(4):225- 
32; Peris et al. Brain Res Mol Brain Res. 1998 Jun 15;57(2):3 10-20; U. S. Patent 

20 5,801,154; U.S. Patent 5,789,573; U. S. Patent 5,718,709 and U.S. Patent 5,610,288). 
Antisense constructs have also been described that inhibit and can be used to tieat a 
variety of abnormal cellular proliferations, e.g. cancer (U. S. Patent 5,747,470; U. S. 
Patent 5,591,317 tod S. Patent 5,783,683). 

Therefore, in certain embodiments, the present invention provides 

25 oligonucleotide sequences that comprise all, or a portion of, any sequence that is 
capable of specifically binding to polynucleotide sequence described herein, or a 
complement thereof. In one embodiment, the antisense oligonucleotides comprise 
DNA or derivatives thereof. In another embodiment, the oligonucleotides comprise 
RNA or derivatives thereof. In a third embodhnent, the oligonucleotides are modified 

30 DNAs comprising a phosphorothioated modified backbone. In a fourth embodiment, 
the oligonucleotide sequences comprise peptide nucleic acids or derivatives diereof* In 



wo 01/96388 



33 



PCT/USOl/18557 



each case, preferred compositions comprise a sequence region that is complementary, 
and more preferably substantially-complementary, and even more preferably, 
completely complementary to one or more portions of polynucleotides disclosed herein* 
Selection of antisense compositions specific for a given gene sequence is based \xpon 

5 analysis of the chosen target sequence and determination of secondary structure, Tm, 
binding energy, and relative stability. Antisense compositions may be selected based 
upon their relative inability to form dimers, hairpins, or other secondary structures that 
would reduce or prohibit specific binding to the target mKNA in a host cell. Highly 
preferred target regions of the mRNA, are fliose which are at or near the AUG 

10 translation uiitiation codon, and those sequences which are substantially complementary 
to 5* regions of the mRNA. These secondary structure analyses and target site selection 
considerations can be performed, for example, using v.4 of the OLIGO primer analysis 
software and/or the BLASTN 2.0.5 algorithm software (Altschul et aU Nucleic Acids 
Res. 1997,25(17):3389-402). 

15 The use of an antisense delivery method employing a short peptide 

vector, termed MPG (27 residues), is also contemplated. The MPG peptide contains a 
hydrophobic domain derived from the fiision sequence of HIV gp41 and a hydrophilic 
domain from the nuclear localization sequence of SV40 T-antigen (Morris et aL, 
Nucleic Acids Res. 1997 Jul 15;25(14):2730-6). It has been demonstrated flwt several 

20 molecules of the MPG peptide coat the antisense oligonucleotides and can be delivered 
into cultured mammalian cells in less than 1 hour with relatively high efficiency (90%). 
Further, the interaction with MPG strongly increases both the stability of the 
oligonuoleotide to nuclease and the ability to cross the plasma membrane. 

According to another embodiment of the invention, the polynucleotide 

25 compositions described hercm are used in the design and preparation of ribozyme 
molecules for inhibiting expression of the tumor polypeptides and protdns of the 
present invention m tumor cells. Ribozymes are KNA-protein complies that cleave 
nucleic adds in a site-specific feshion. Ribozymes have specific catalytic domains that 
possess endonuclease activity (Kim and Cech, Proc Nafl Acad Sd U S A, 1987 

30 Dec;84(24):8788-92; Forster and Symons, Cell. 1987 Apr 24;49(2):211-20). For 
example, a large number of ribozymes accelerate phosphoester transfix reactions with a 



BNSDOCID: <WO^__01fl838aA?_L> 



wo 01/96388 



34 



PCTAUSOl/18557 



high degree of specificity, often cleaving only one of several phosphoesters in an 
oligonucleotide substrate (Cech et d.. Cell. 1981 Dec;27(3 Pt 2):487-96; Michel and 
Westhot J Mol Biol. 1990 Dec 5;216(3):585-610; Reinhold-Hurek and Shub, Nature. 
1992 May 14;357(6374):173-6). This specificity has been attributed to the requirement 

5 that the substrate bind via specific base-paiiing interactions to the internal guide 
sequence ("IGS") of the ribozyme prior to chemical reaction. 

Six basic varieties of naturally-occurring ©ozymatic KNAs are known 
presently. Each can catalyze the hydrolysis of KNA phosphodiester bonds in trans (and 
thus can cleave other RNA molecules) under physiological conditions. In general, 

10 enzymatic nucleic acids act by first binding to a target RNA. Such binding occurs 
through the target binding portion of a enzymatic nucleic acid which is held in close 
proximity to an enzymatic portion of the molecule that acts to cleave the target RNA. 
Thus, the en2ymatic nucleic acid first recognizes and then binds a target KNA through 
complementary base-pairing, and once bound to the correct site, acts ensymatically to 

IS cut the target RNA. Strategic cleavage of sudi a target RNA will destroy its ability to 
direct synthesis of an encoded protein. After an enzymatic nucleic add has bound and 
cleaved its RNA target, it is released ftom that RNA to search for another target and can 
repeatedly bind and cleave new targets. 

The enzymatic nature of a ribozyme is advantageous over many 

20 technologies, such as antisense technology (where a nucleic acid molecule simply binds 
to a nucleic add target to block its translation) since Ae concratration of ribozyme 
necessary to afiFect a therapeutic treatment is lower than that of an antisense 
oligonucleotide. ' This advantage reflects the ability of the ribozyme to act 
enzymatically. Thus, a single ribozyme molecule is able to cleave many molecules of 

25 target RNA. In addition, the ribozyme is a highly specific inhibitor, with the ^ecifidty 
of inhibition depending not only on the base pairing mechaiiism of binding to the target 
RNA, but also on the mechanism of target RNA cleavage. Smgle mismatches, or base- 
substitutions, near the site of cleavage can completely eliminate catalytic activity of a 
ribozyme. Similar mismatches in antisense molecules do not prevent their action 

30 (Woolf et al, Proc NaU Acad Sci U S A. 1992 Aug lS;89(16):7305-9). Thus, the 



BNSOOCID: <W0 ^0186388A2JL> 



wo 01/96388 PCT/USOl/18557 

35 

specificity of action of a nbo^mie is greater than (hat of an antisense oligonucleotide 
binding the same RNA site. 

The enzymatic nucleic acid molecule may be fonned in a hammeibead, 
hairpin, a hepatitis 5 virus, group I intton or RNaseP RNA (in association wilb an RNA 

5 guide sequence) or Neurospora VS RNA motif Examples of hammeihead motifs are 
described by Rossi et al Nucleic Acids Res. 1992 Sep 1 1;20(17):4559-65. Examples of 
hairpin motifs are described by Hampel et al (Eur, Pat. Appl. Publ. No. EP 0360257), 
Hampel and Tritz, Biochemistry 1989 Jun 13;28(12):4929-33; Hampel etaL, Nucleic 
Acids Res. 1990 Jan 25;18(2):299-304 and U. S. Patent 5,631,359. An example of the 

10 hepatitis S virus motif is described by Perrotta and Been, Biochemistry. 1992 Dec 
1;31(47):1 1843-52; an example of the RNaseP motif is described by Guerrier-Takada 
era/.. Cell. 1983 Dec;35(3 Pt 2):849-57; Neurospora VS RNA ribozyme motif is 
described by Collins (Saville and CoDins, Cell. 1990 May 18;61(4):685-96; Saville and 
Collins, Proc Natl Acad Sci U S A, 1991 Oct l;88(19):8826-30; Collins and Olive, 

15 Biochemistry. 1993 Mar 23;32(1 1):2795-9); and an example of the Group I intron is 
described in (U. S. Patent 4,987,071). All diat is important in an enzymatic nucleic acid 
molecule of this invention is that it has a specific substrate binding site which is 
complementary to one or more of the target gene RNA regions, and that it have 
nucleotide sequences within or surrounding that substrate binding site which impart an 

20 RNA cleaving activity to the molecule. Thus the ribozyme constructs need not be 
limited to specific motifs mentioned herein. 

Riboqmes may be designed as described in Int. Pat Appl. PubL No. 
WO 93/23569 and Int. Pat Appl. PubL No. WO 94/02595, each specifically 
incorporated herem by reference) and synthesized to be tested in vitro and in vfvo, as 

25 described. Such ribosymes can also be optimized for delivery. While specific 
examples are provided, those in the art will recognize that equivalent RNA targets in 
other species can be utilized ^en necessary. 

Ribozyme activity can be optimized by altering the length of the 
ribozyme binding arms, or chemically synthesizing ribozymes with modifications that 

30 prevent their degradation by serum ribonucleases (see e.g., Int Pat. Appl. Publ. No. 
WO 92/07065; Int.Pat Appl. Publ. No. WO 93/15187; hit Pat Appl. Publ. No. WO 



BNSOOCID: ^WO ^0196388A^l^ 



wo 01/96388 



36 



PCT/US01/18S57 



91/03162; Eur. Pat AppL Publ. No. 92110298,4; U. S. Patent 5,334,711; and Int. Pat 
Appl. Publ. No. WO 94/13688, which describe various chexnical modifications that can 
be made to the sugar moieties of enzymatic KNA molecules), modifications which 
enhance their efiScacy in cells, and removal of stem U bases to shorten RNA synthesis 

5 times and reduce chemical requirements. 

Sullivan etal. (tot Pat Appl. Publ. No. WO 94/02595) describes the 
general methods for delivery of enzymatic RNA molecules. Ribozymes may be 
administered to cells by a variety of methods known to those familiar to the art, 
including, but not restricted to, encapsulation in liposomes, by iontophoresis, or by 

10 incorporation into other vehicles, such as hydrogels, cyclodextrins, biodegradable 
nanocapsules, and bioadhesive microspheres. For some indications, ribozymes may be 
directly delivered ex vivo to ceUs or tissues Avith or without the aforementioned 
vehicles. Alternatively, the RNA/vehicle combination may be locally delivered by 
direct inhalation, by direct, injection or by use of a catheter, infusion pump or stent 

15 Other routes of delivery include, but are not limited to, intravascular, mtramuscular, 
subcutaneous or joint injection, aerosol inhalation, oral (tablet or pill foim), topical, 
systemic, ocular, intraperitoneal and/or mtrathecal deliyeiy^ More detailed descriptions 
of ribozyme delivery and administration are provided in Int Pat Appl. Publ. No. WO 
94/02595 and tot. Pat Appl. Publ. No.' WO 93/23569, each specifically mcorporated 

20 hereto by reference. 

Another means of accumulatmg high concentrations of a ribozyme(s) 
within cells is to incorporate the ribozyme-encoding sequences mto a DNA expression 
vector. Transcription of the ribozyme sequences are driven from a promoter for 
eukaryotic RNA polymerase I (pol I), RNA polymerase II (pol II), or RNA polymerase 

25 III (pol III). Transcripts from pol II or pol III promoters will be expressed at high levels 
in all cells; the levels of a given pol II promoter in a given cell type will depend on the 
nature of the gene regulatory sequences (enhancers, silencers, etc.) present nearby. 
Prokaryotic RNA polymerase promoters may also be used, providtog that flie 
prokaryotic RNA polymerase enzyme is ejqpressed to the appropriate cells Ribozymes 

30 expressed from such promoters have been shown to function to mammalian cells. Such 
transcription units can be tocozporated toto a variety of vectors for totroduction toto 



BNS00CID:<WO. 



I196388ASJU» 



wo 01/96388 



PCTAJSOl/18557 



37 

mammalian cells, including but not restricted to, plasmid DNA vectors, viral DNA 
vectors (such as adenovirus or adeno-associated vectorsX or viral RNA vectors (such as 
letroviial, semliki forest virus, sindbis virus vectors). 

In ano&er embodiment of ihe invention, peptide jnicleic acids (FNAs) 

5 compositions are provided. PNA is a DNA mimic in vAAdx flie nucleobases are 
attached to a pseudopeptide backbone (Good and Nielsen, Antisense Nucleic Acid Drug 
Dev. 1997 7(4) 431-37). PNA is able to be utilized in a number methods that 
traditionally have used RNA or DNA. Often PNA sequences perform better in 
techniques than the corresponding RNA or DNA sequences and have utilities that are 

10 not inherent to RNA or DNA, A review of PNA including methods of making, 
characteristics of, and methods of usmg, is provided by Corey (TVemZy Biotechnol 1997 
Jun;15(6):224-9). As such, in certain embodiments, one may prepare PNA. sequences 
that are complementary to one or more portions of the ACE mRNA sequence, and sucsh 
PNA compositions may be used to regulate, aher, decrease, or reduce ihe translation of 

15 ACE-specific mRNA, and thereby alter the level of ACE activity in a host cell to v^^ich 

such PNA compositions have been administered. 

PNAs have 2-anunoethyl-glycine linkages repladng the normal 

phosphodiester backbone of DNA (Nielsen et a/., Sdmce 1991 Dec 6^54(5037):1497- 

500; Hanvey et aU Science. 1992 Nov 27;258(5087):1481-5; Hyrup and Nielsen, 

20 Bioorg Med Chem. 1996 Jan;4(l):5-23). This chemistry has toee important 
consequences: firstly, in contrast to DNA or phosphorotfaioate oligonucleotides, PNAs 
arc neutral molecules; secondly, PNAs are achhal, which avoids the need to develop a 
stereoselective synthesis; and thirdly, PNA synthesis uses standard Boc or Fmoc 
protocols* for solid-phase pqptide syntiiesis, although othor methods, including a 

'25 modified Merrifield method, have been used. 

PNA monomers or ready-made oligomers are commercially available 
firom PerSeptive Biosystems (Framingham, MA). PNA syntheses by either Boc or 
Fmoc protocols are straightforward using manual or automated protocols (Norton et al, 
Bioorg Med Chem. 1995 Apr;3(4):437-45). The manual protocol lends itself to the 

30 production of chemically modified PNAs or the simultaneous synthesis of families of 
closely related PNAs. 



BN8D0CI0: **«>__Ot9638aWUJ> 



wo 01/96388 PCTAJS01/J85S7 

38 

As with peptide synthesis, the success of a particular PNA synthesis will 
depend on the properties of the chosen sequence. For example, while in theory PNAs 
can incorporate any combination of nucleotide bases, the presence of adjacent purines 
can lead to deletions of one or more residues in the product, hi expectation of this 
5 difficulty, it is suggested that, in producing PNAs with adjacent purines, one should 
repeat the couplmg of residues likely to be added ineflSciently. This should be followed 
by the purijScation of PNAs by reverse-plmse high-pressure liquid chromatography, 
providing yields and purity of product similar to those observed during the synthesis of 
pq>tides. 

10 Modifications of PNAs for a given application may be accomplished by 

coiq)ling amino acids during solid-phase synthesis or by attaching compounds that 
contain a carboxylic acid group to the exposed N-terminal amine. Ahematively, PNAs 
can be modified after synlfaeas by coirpling to an introduced lysine cysteme. The 
ease with which PNAs can be modified fedlitates optimization for better solubility or 

IS for spedfic functional requiremmts. Qncesynthedzed, the identity of PNAs and their 
derivatives can be confirmed by mass spectrometry. Several studies have made and 
utilized modifications of PNAs (for example, Norton et aL, Biooi^ Med Chem. 199S 
Apr;3(4):437-45; Petersen et cd., J Pept Sd. 1995 May-Jun;l(3):175-83; Oram et d., 
Biotechniques. 1995 Sep;19(3):472-80; Footer et aL, Biochemistry. 1996 Aug 

20 20^J5(33):10673-9; Griffith et aL, Nucleic Adds Res. 1995 Aug 11^3(15):3003-8; 
Pardridge et aL, Proc Natl Acad Sd U S A. 1995 Jun 6;92(12):5592-6; Bofifa et aL, 
Proc Natl Acad Sci USA. 1995 Mar 14;92(6):1901-5; Gambacorti-Passerini et aL, 
Blood. 1996 Aug 15;88(4):1411-7; Armitage et al., Proc Nafl Acad Sd USA. 1997 
Nov ll;94(23):12320-5; Seeger et aL, Biotechniques. 1997 Sep;23(3):512-7). U.S. 

25 Patent No. 5,700,922 discusses PNA-DNA-PNA chuneric molecules and their uses in 
diagnostics, modulating protein in organisms, and treatment of conditions susceptible to 
therapeutics. * 

Methods of characterizing the antisense binding properties of PNAs are 
discussed in Rose (Anal Chem. 1993 Dec 15;65(24):3545-9) and Jensen et aL 
30 (Biochemistry. 1997 Apr 22;36(16):5072-7). Rose uses capillary gel electrophoresis to 
determine binding of PNAs to their complementary oligonucleotide, measuring the 
relative binding kinetics and stoichiometry. Similar types of measurements were made 
by Jensen et aL using BIAcore™ technology. 



BNSDOQD: <W0 0196388A2J_> 



wo 01/96388 PCTAIS01/18S57 

39 

Other applications of PNAs that have been described and will be 
apparent to the skilled artisan include use in DNA strand invasion, antisense inhibition, 
mutational analysis, enhancers of transcription, nucleic acid purification, isolation of 
transcriptionally active genes, blocking of transcription factor binding, genome 
5 cleavage, biosensors, in situ hybridization, and the like. 

Polynucleotide Identification, Characterization and Expression 

Polynucleotides compositions of the present invention may be identified, 
prepared and/or manipulated using any of a variety of well established techniques (see 
generally, Sambrook et al., Molecular Cloning: A Laboratory Manual, Cold Spring 

10 Harbor Laboratories, Cold Spring Harbor, NY, 1989, and other like ref^ences). For 
example, a polynucleotide may be identified, as described in more detail below, by 
screening a microarray of cDNAs for tumor-associated expression {i.e., expression that 
is at least two fold greater in a tumor than in normal tissue, as det^xoined using a 
representative assay provided herein). Such sheens may be poformed, for example, 

15 using the microarray technology of Afifymetrix, Inc. (Santa Clara, CA) according to the 
manu&cturer's instructions (and essentially as described by Schena et al., Proc. Natl 
Acad. ScL USA P3:10614-10619, 1996 and Heller et al., Proc. Natl Acad Sci. USA 
W:2150-2155, 1997). Alternatively, polynucleotides may be amplified fiom cDNA 
prepared from cells expressing the proteins described herein, such as tumor cells. 

20 Many tanplate dependent processes are available to amplify a target 

sequences of interest present in a sample. One of the best known amplification methods 
is the polymerase bhain reaction (PGR™) which is described in detail in U.S. Patent 
Nos. 4,683,195, 4,683,202 and 4,800,159, each of Which is incorporated herem by 
reference in its entirety. Briefly, in PGR™, two primer sequences are prepared which 

23 are complementary to regions on opposite complementary strands of the target 
sequence. An excess of deoxynucleoside triphosphates is added to a reaction mixture 
along with a DNA polymerase (e.g., Tag polymerase). If the target sequence is present 
in a sample, the primers will bind to Ae target and the polymerase will cause the 
primers to be extended along the target sequence by adding on nucleotides. By raising 

50 and lowering the temperature of the reaction mixture, the extended primers will 



BtlSDOCiO: <WW)__019e38aAa»L^ 



wo 01/96388 



40 



PCTAJS01/18S57 



dissociate from the target to forai reaction products, excess primers will bind to the 
target and to the reaction product and the process is repeated. Preferably reverse 
transcription and PCR™ amplification procedure may be performed in order to quantify 
the amount of mKNA amplified. Polymerase cham reaction i^ieihodologies are well 

5 known in the art. 

Any of a number of other template dependent processes, many of which 
are variations of the PGR ™ amplification technique, are readily known and available in 
the art. Illustratively, some such methods include the ligase chain reaction (referred to 
as LCR), described, for example, in Eur. Pat. Appl. Publ. No. 320,308 and U.S. Patent 

10 No. 4,883,750; Qbeta Replicase, described in PCX Intl. Pat. Appl. Publ. No. 
PCTAJS87/00880; Strand Displacement Amplification (SDA) and Repaff Chain 
Reaction (RCR). Still other amplification methods are described in Great Britain Pat 
Appl. No. 2 202 328, and in PCT totl. Pat. Appl. Publ, No. PCTAJS89/01025. Other 
nucleic acid amplification procedures include transcription-based amplification systems 

15 (TAS) (PCT Ml Pat Appl. Publ. No. WO 88/10315), including nucleic acid sequence 
based amplification (NASBA) and 3SR. Eur. Pat. Appl. Publ. No. 329,822 describes a 
nucleic add amplification process involving cyclically synthesizing single-stranded 
RNA ("ssRNA"), ssDNA, and doubl^stranded DNA (dsDNA). PCT Intl. Pat. Appl. 
Publ. No. WO 89/06700 describes a nucleic acid sequence amplification scheme based 

20 on the hybridization of a promoter/primer sequence to a target single-stranded DNA 
("ssDNA") followed by transcription of many RNA copies of the sequence. Other 
amplification methods such as "RACE" (Frohman, 1990), and •'one-sided PGR" (Ohara, 
1989) are also well-known to those of skill in the art. 

An amplified portion of a polynucleotide of the present invention may be 

25 used to isolate a fiill length gene from a suitable library (e.g., a tumor cDNA library) 
using well known techniques. Within such techniques, a library (cDNA or genomic) is 
screened using one or more polynucleotide probes or primers suitable for amplification. 
Preferably, a library is size-selected to include larger molecules. Random primed 
libraries may also be preferred for identifying 5' and upstream regions of genes. 

30 Genomic libraries are preferred for obtaining introns and extending 5' sequences. 



BNSDOCID: <WO Oig6a88A?^L> 



wo 01/96388 



41 



PCT/US01/18S57 



For hybridization techniques, a partial sequence may be labeled (e.g., by 
nick-translation or end-labeling with ^^P) using well-known techniques. A bacterial or 
bacteriophage library is then generally screened by hybridiang filters containing 
denatured bacterial colonies (or lawns contaimng phage plaques)^th the labeled probe 

5 (see Sambrook et al.. Molecular Cloning: A Laboratory Manual Cold Spring Harbor 
Laboratories, Cold Spring Harbor, NY, 1989). Hybridizing colonies or plaqiK^s are 
selected and expanded, and the DNA is isolated for further analysis. cDNA clones may 
be analyzed to determine the amount of additional sequence by, for example, PCR 
using a primer from the partial sequence and a primer from the vector. Restriction 

10 maps and partial sequences may be generated to identify one or more overlappmg 
clones. The complete sequence may then be determined usii% standard techniques, 
which may involve generating a series of deletion clones. The resulting overlapping 
sequences can then assembled into a single contiguous sequence. A friU length cDNA 
molecule can be generated by ligating suitable fi-agments, using well known techniques. 

15 Alternatively, amplification techniques, such as those described above, 

can be usefiil for obtaining a fiill length coding sequence fix)m a partial cDNA 
sequence. One such amplification technique is inverse PCR {see Triglia et al., NucL 
Acids Res, /tf:8186, 1988), which uses restriction enzymes to generate a fragment in the 
known region of the gene. The fragment is then circularized by intramolecular ligation 

20 and used as a template for PCR with divergent primers derived from the known region. 
Within an alternative approach, sequences adjacent to a partial sequence may be 
retrieved by amplification with a primer to a linker sequence and a primer specific to a 
known region. The amplified sequences are typically subjected to a second round of 
amplification with flie same linker primer and a second primer specific to the known 

25 region. A variation on tiiis procedure, which OTiploys two primers tiiat initiate 
extension in opposite directions from the known sequence, is described in WO 
96/38591. Another such technique is known as "r^id amplification of cDNA ends" or 
RACE. This technique involves the use of an internal primer and an external primer, 
vdiich hybridizes to a polyA region or vector sequence, to identify sequences that are 5* 

30 and 3' of a known sequence. Additional techniques include capture PCR (Lagerstrom et 
PCR Methods Applic, i:llM9, 1991) and walking PCR (Parker et al., iVwc/. Acids. 



BMSooao: <wo___oi96ai»AaLL> 



wo 01/96388 



42 



PCT/USOl/18557 



Res. 7P:3055-60, 1991)- Other methods employing amplification may also be 
^ployed to obtain a &11 length cDN A sequence. 

In certain instances, it is possible to obtain a full laigth cDNA sequence 
by analysis of sequences provided in an expressed sequence tag ^BST) database^ such as 

5 that available fiom GenBank. Searches for overlapping ESTs may generally be 
performed using well known programs {e.g., NCBI BLAST searches), and such ESTs 
may be used to generate a contiguous full length sequence. Full length DNA sequences 
may also be obtained by analysis of genomic fragments. 

In other embodiments of the invention, polynucleotide sequences or 

10 fragments thereof which encode polypeptides of the invention, or fusion proteins or 
functional equivalents th^of, may be used in recombinant DNA molecules to direct 
expression of a polypeptide m appropriate host cells. Due to the inherent degeneracy of 
the genetic code, other DNA sequences that encode substantially the same or a 
functionally equivalent amino acid sequence may be produced and these sequences may 

IS be used to clone and express a given polypeptide. 

As will be understood by those of skill in the art, it may be advantageous 
in some instances to produce polypeptide-encoding nucleotide sequences possessing 
non-naturally occurring codons. For example, codons preferred by a particular 
prokaryotic or eukaryotic host can be selected to increase the rate of protein expression 

20 or to produce a recombinant RNA transcript having desirable properties, such as a half- 
life which is longer than that of a transcript generated from the naturally occurring 
sequence. 

Mbreover, the polynucleotide sequences of the present invention can be 
engineered using methods generally known in the art in order to alter polypeptide 

25 encoding sequences for a variety of reasons, including but not limited to, alterations 
which modify the cloning, processing, and/or expression of the gene product For 
example, DNA shuffling by random fragmentation and PGR reassembly of gene 
fragments and synthetic oligonucleotides may be used to engineer the nucleotide 
sequences. In addition, site*directed mutagenesis may be used to insert new restriction 

30 sites, alter glycosylation pattans, change codon preferwxce, produce splice variants, or 
introduce mutations, and so forth. 



BNSDOaO: <Vro ^01fl6388A%JL> 



wo 01/96588 PCT/USOl/18557 

43 

In anoth^ embodimrat of the inventioii, natural, modified, or 
recombinant nucleic acid sequences may be ligated to a heterologous seqiience to 
encode a fusion protein. For example, to screen peptide libraries for inhibitors of 
polypeptide activity, it may be useful to encode a chimeii^rotem that can be 

S recognized by a commercially available antibody. A fusion protein may also be 
engineered to contain a cleavage site located between the polypeptide-encoding 
sequence and the heterologous protein sequence, so that the polypeptide may be cleaved 
and purified away firom the heterologous moiety. 

Sequences encoding a desired polypeptide may be synthesized, in whole 

10 or in part, usmg^ chemical methods well known in the art (see Caruthors, M. H. et al. 
(1980) Nucl Acids Res. Synq^. Ser. 215-223, Horn, T. et al. (1980) NucL Acids Res. 
Symp, Sen 225-232). Alternatively, the protein itself may be produced uang chemical 
methods to synfliesize tibie amino add sequence of a polypeptide, or a portion thereof. 
' For example, peptide synthesis can be performed using various solid-phase techniques 

15 (Rob^ge, J. Y. et al. (1995) Science 269:202-2M) and automated synthesis may be 
achieved, for example, using the ABI 431 A Peptide Synthesizer (Peiidn Ehner, Palo 
Alto, CA). 

A newly synthesized peptide may be substantially purified by 
preparative high performance liquid chromatography (e.g., Creighton, T. (1983) 

20 Proteins, Structures and Molecular Principles, WH Freeman and Co., New York, N.Y.) 
or other comparable techniques available in the art. The composition of the synthetic 
peptides may be confirmed by amino acid analysis or sequencing (e.g., the Edman 
degradation procedure). Additionally, the amino add sequence of a polypeptide, or any 
part th^eof, may be altered during direct synthesis and/or combined udng chemical 

25 methods with sequences from other protems, or any part fliereof, to produce a variant 
polypeptide. 

In order to express a desired polypqitide, the nucleotide sequences 
encodmg the polypeptide, or functional equivalents, may be inserted mto appropriate 
expression vector, a vector which contains the necessary elements for the 
30 transcription and translation of the inserted coding sequence. Methods which are well 
known to those skilled in the art may be used to construct expression vectors containing 



BNSDOCID: <W0 ^019638aA?_L> 



wo 01/96388 PCT/US01/J8557 

44 

sequences encoding a polypeptide of interest and appropriate transcriptional and 
translational control elements. These methods include in vitro recombinant DNA 
techniques, synthetic techniques, and in vivo genetic recombinatioa Such techniques 
are described, for example, in Sambrook, J. et al. (1989) Molecular Cloning, A 
5 Laboratory Manual, Cold Spring Harbor Press, Plainview, N.Y., and Ausubel, F. M. et 
al. (1989) Current Protocols in Molecular Biology, John Wiley & Sons, New Yoric. 
N.Y. 

A variety of expression vector/host systems may be utilized to contam 
and express polynucleotide sequences. These include, but are not limited to, 

1 0 microorganisms such as bacteria transformed with recombinant bacteriophage, plasmid, 
or cosmid DNA expression vectors; yeast transformed with yeast expression vectors; 
insect cell systems infected with virus expression vectors {e.g.y baculovirus); plant cell 
systems transformed with virus expression vectors (e.g., cauliflower mosaic virus, 
CaMV; tobacco mosaic virus, TMV) or with bacterial expression vectors (e.g., Ti or 

1 S pBR322 plasmids); or animal cell systems. 

The "control elements" or "regulatory sequences" present m an 
expression vector are those non-translated regions of the vector-enhancers, promoters, 
S* and 3* untranslated regions-which interact with host cellular proteins to carry out 
transcription and translation. Such elements may vary in their strength and specifidty. 

20 Depending on the vector system and host utilized, any number of suitable transcription 
and translation elemente, inchiding constitutive and inducible promoters, may be used. 
For example, when cloning in bacterial systems, inducible promoters such as the hybrid 
lacZ promoter of the pBLUESCRIPT phagemid (Stratagene, La JoUa, Calif.) or 
pSPORTl plasmid (Gibco BRL, Gaithersburg, MD) and the like may be used. In 

25 nuuiunalian cell systems, promoters from mammalian genes or from mammalian 
viruses are generally preferred. If it is necessary to generate a cell line that contains 
multiple copies of the sequence encoding a polypeptide, vectors based on S V40 or EBV 
may be advantageously used with an appropriate selectable marker. 

In bactOTal systems, any of a number of expression vectors may be 

30 selected depending upon the use intended for the expressed polypeptide. For example, 
when large quantities are needed, for example for the induction of antibodies, vectors 



BNS00CID:^WO ^01B6388ASLJL> 

_ _ , ... • •."•"-•"r.'Wtsvwxfi 



wo 01/96388 



PCTAJSOl/18557 



45 

vAdch direct high level expression of fusion proteins that are readily purified may be 
used. Such vectors include, but are not limited to, the multiiunctional K coli cloning 
and expression vectors such as pBLUESCRIPT (Stratagrae), in which the sequence 
encoding the polypeptide of interest may be ligated into the A^ector in fiame with 

S sequraces for the amino-temiinal Met and the subsequrat 7 residues of .beta.- 
galactosidase so that a hybrid protein is produced; pIN vectors (Van Heeke, G. and S. 
M. Schuster (1989) 1 BioL Chem. 264:5503-5509); and the like. pGEX Vectors 
(Promega, Madison, Wis.) may also be used to express foreign polypeptides as fusion 
proteins with glutathione S-transferase (GST). In general, such fusion proteins are 

10 soluble and can easily be purified fiom lysed cells by adsorption to glutathione-agarose 
beads followed by elution in the pxeseDce of free glutathione. Proteins made in such 
systems may be designed to include heparin, thrombin, or factor XA protease cleavage 
sites so that the clon€»d polypeptide of interest can be released from the GST moiety at 
will. 

IS In the yeast, Saccharomyces cerevisiae, a number of vectors containing 

constitutive or inducible promoters such as alpha factor, alcohol oxidase, and PGH may 
be used. For reviews, see Ausubel et al. (supra) and Grant et al. (1987) Methods 
Enzymol 153:516-544. 

In cases where plant expression vectors are used, the expression of 

20 sequences encoding polypeptides may be driven by any of a number of promoters. For 
example, viral promoters such as the 35S and 19S promoters of CaMV may be used 
alone or in combination with the omega leader sequence from TMV (Takamatsu, N. 
(1987) MMBO J. tf:307-3 11. Alternatively, plant promoters such as the small subunit of 
RUBISCO or heat shock promoters may be used (Conizzi, G. et al. (1984) EMBO 1 

25 3:1671-1680; Broglie, R. et al. (1984) Science 22^:838-843; and Whiter, J. et al. (1991) 
Results Probl Cell Differ, 77:85-105). These constructs can be introduced into plant 
cells by direct DNA transformation or pathogen-mediated transfection. Such techniques 
are described in a number of generally available reviews (see, for example, Hobbs, S. or 
Muny, L. E. in McGraw HBU Yearbook of Scimce and Technology (1992) McGraw 

30 Hill, New York, N.Y.; pp. 191-196). 



i9Ga8aA%JL> 



wo 01/96388 



46 



PCTAJSOl/18557 



An insect system may also be used to express a polypeptide of interest. 
For example, in one such system, Autographa califomica nuclear polyhedrosis virus 
(AcNPV) is used as a vector to express foreign genes in Spodoptera fiugiperda cells or 
in Trichoplusia larvae. The sequences encoding the polypeptide jnay be cloned into a 

S non-essential region of the virus, such as the polyhedrin gene, md placed under control 
of the polyhedrin promoter. Successful insertion of the polypeptide-encoduig sequence 
will render the polyhedrin gene inactive and produce recombinant virus lacking coat 
protein. The recombinant viruses may then be used to infect, for example, S. firugiperda 
cells or Trichoplusia larvae in which the polypeptide of interest may be expressed 

10 (Engelhard, E. et al. (1994) Proc. Natl Acad Sci 91 :3224-3227). 

In mammalian host cells, a number of viral-based expression systems are 
generally available. For example, in cases where an adenovirus is used as an expression 
vector, sequences encodir^ a polypeptide of interest may be ligated into an adenovirus 
transcription/translation complex consisting of the late promoter and tripartite leader 

15 sequence. Insertion in a non-essential El or E3 region of the viral genome may be used 
to obtain a viable virus which is capable of expressing the polypeptide in infected host 
cells (Logan, J. and Shenk, T. (1984) Proc. Natl Acad Sci 81:3655-3659). In addition, 
transcription enhancers, such as the Rous sarcoma virus (RSV) enhancer, may be used 
to increase expression in manmialian host cells. 

20 Specific initiation signals may also be used to achieve more efficient 

translation of sequences encoding a polypeptide of interest. Such signals include Ae 
ATG faiitiation codon and adjacent sequences. In cases where sequences encoding the 
polypeptide, its initiation codon, and upstream sequences are inserted into the 
appropriate expression vector, no additional transcriptional or translational control 

25 signals may be needed. However, in cases where only coding sequence, or a portion 
thereof, is inserted, exogenous translational control signals including the ATG initiation 
codon should be i»:ovided. Furthermore, the initiation codon should be in the correct 
reading firame to ensure translation of the mtire insert. Exogenous translational 
elements and initiation codons may be of various origins, both natural and synthetic. 

30 The efiBciracy of expression may be enhanced by the inclusion of enhancers which are 



BNSOOGID: 



wo 01/96388 



47 



PCT/USOl/18557 



appropriate for the particular cell system which is used, such as those described in the 
literature (Scharf, D. et al. (1994) Results Probl Cell Differ. 20:125-162). 

r 

In addition, a host cell strain may be chosm for its ability to modulate 
the expression of tiie inserted sequences or to process the expressed protein in the 

5 desired fashion. Such modifications of the polypeptide includCr but are not limited to, 
acetylation, carboxylation. glycosylation, phosphorylation, lipidation, and acylation. 
Post-translational processing which cleaves a "prepro" form of the protein may also be 
used to facilitate correct insertion, folding and/or function. Different host cells such as 
CHO, COS, HeLa, MDCK, HEK293, and WI38, which have specific cellular 

10 machinery and characteristic mechanisms for such post-translational activities, may be 
chosen to ensure the correct modification and processing of the foreign protein. 

For long-term, high-yield production of recombinant proteins, stable 
expression is generally preferred. For ocample, cell lines which stably express a 
polynucleotide of interest may be transformed usmg expression vectors which may 

IS contain viral origins of r^lication and/or endogenous exi»ession elements and a 
selectable marker gene on the same or on a separate vector. Following &e introduction 
of the vector, cells may be allowed to grow for 1-2 days in an enriched media before 
they are switched to selective media. The purpose of the selectable marker is to confer 
resistance to selection, and its presence allows growth and recovery of cells which 

20 successfully express tiie introduced sequences. Resistant clones of stably transformed 
cells may be proliferated usmg tissue culture techniques appropriate to the cell type. 

Any nxmiber of selection systems may be used to recover transformed 
cell lines. These include, but are not limited to, the herpes sunplex virus fliymidine 
kinase (Wigler, M. et al. (1977) Cell 77:223-32) and adenine phosphoribosyltransferase 

25 (Lowy, 1. et al. (1990) Cell 22:817-23) genes which can be employed in tk.sup.- or 
aprt.sup.- cells, respectively. Also, antimetabolite, antibiotic or herbicide resistance can 
be used as the basis for selection; for example, dhfi: whidi confers resistance to 
methotrexate (Wigler, M. et al. (1980) Proc. Natl. Acad ScL 77:3567-70); npt, which 
' confers resistance to the aminoglycosides, neomycin and G-418 (Colberc-Gar^in, F. et 

30 al (1981) y. Mol Biol 150:M4); and als or pat, which confer resistance to 
chlorsulfiiron and phosphinotricin acetyltransferase, respectively (Muny, ^wpra). 



BN800CID: <VW>__018638aA^JL> 



wo 01/96388 



48 



PCTAJSOl/18557 



Additional selectable genes have been described, for example, ftpB, which allows cells 
to utilize indole in place of tryptophan, or hisD, vAich allows cells to utilize histinol in 
place of histiduie (Hartman, S. C. and R. C. Mulligan (1988) Proc. NatL Acad ScL 
SJ:8047-51X The use of visible markers has gained popularitywith sudi markers as 
5 anthocyamns, beta-glucuronidase and its substrate GUS, and luciferase and its substrate 
luciferin, being widely used not only to identify transformants, but also to quantify the 
amount of transient or stable protein expression attributable to a specific vector system 
. (Rhodes, C. A. et al. (1995) Methods MoL Biol 55:121-131). 

Although the presence/absence of marker gene expression suggests that 
10 the gene of interest is also present, its presence and expression may need to be 
confirmed- For example, if the sequence encoding a polypeptide is inserted within a 
marker gene sequence, recombinant cells containmg sequences can be identified by the 
absmce of marker gene fimction. Alternatively, a marker gene can be placed in tandem 
with a polypeptide-encoding sequence under the control of a single promoter. 
15 Expression of tiie marker gene in response to mduction or selection usually indicates 
expression of tiie tandem gene as well. 

Alternatively, host cells that contain and express a desired 
polynucleotide sequence may be identified by a variety of procedures known to fliose of 
skill in the art These procedures include, but are not limited to, DNA-DNA or DNA- 
20 RNA hybridizations and protein bioassay or immunoassfQf techniques v*ich include, 
for example, membrane, solution, or chip based technolo&es for tiie detection and/or 
quantification of nucleic acid or protein. 

A variety of protocols for detecting and measuring the expression of 
polynucleotide-encoded products, using either polyclonal or monoclonal antibodies 
25 specific for flie product are known in the art. Examples include enzyme-linked 
immunosorbent assay (ELISA), radiounmunoassay (RIA), and fluorescence activated 
cell sorting (FACS). A two-site, monoclonal-based unmunoassay utilizing monoclonal 
antibodies reactive to two non-interfering epitopes on a given polypeptide may be 
preferred for some applications, but a competitive binding assay may also be employed. 
30 These and other assays are described, among other places, in Hampton, R. et al. (1990; 



BNSOOCID: ^__0196388A«_L^ 



wo 01/96388 



49 



PCTA)S01/18557 



Serological Mefibods, a Laboratory Manual, APS Press, St Paul. Minn.) and Maddox, D. 
E, et al. (1983; J. Exp, Med /J5:121 1-1216). 

A wide variety of labels and conjugation techniques are known by those 
skilled in the art and may be used in various nudeic acid and amino add assays. Means 

5 for producing labeled hybridization or PGR probes for detecting sequences related to 
polynucleotides include oligolabeling, nick translation, end-labeling or PGR 
amplification using a labeled nucleotide. Alternatively, the sequences, or any portions 
thereof may be cloned into a vector for the production of an mRNA probe. Such vectors 
are known in the art, are commercially available, and may be used to synthesize RNA 

1 0 probes in vitro by addition of an appropriate KNA polymerase such as T7, T3, or SP6 
and labeled nucleotides. These procedures may be conducted using a variety of 
commercially available kits. Suitable reporter molecules or labds, which may be used 
include radionuclides, enqrmes, fluorescent, diemiluminescent, or chromogenic agents 
as well as substrates, cofactors, inhibitors, magnetic particles, and the like. 

1 5 Host cells transformed with a polynucleotide sequence of interest may be 

cultured under conditions suitable for tiie expression and recovery of the protein from 
cell cuttare. The protdn produced by a recombinant cell may be secreted or contained 
intracellularly depending on the sequence and/or the vector used. As will be understood 
by those of skill in the art, expression vectors containing polynucleotides of the 

20 invention may be designed to contain signal sequences which direct secretion of the 
encoded polypeptide through a prokaryotic or eukaryotic cell membrane. Other 
recombinant constructions may be used to join sequences encoding a polypeptide of 
interest to nucleotide sequence encoding a polypeptide domain which will facilitate 
purification of soluble proteins. Sudi purification facilitatmg domains mclude, but are 

25 not limited to, metal chelating pqstides such as histidine-tryptophan modules that allow 
purification on immobilized metals, protein A domains that allow purification on 
immobilized immunoglobulin, and the domain utilized in the FLAGS extension/a£5nity 
purification system (Immunex Coip., Seattle, Wash.). The indusion of cleavable linker 
sequences such as those specific for Factor XA or enterokinase (Invitrogen. San Diego, 

30 Calif.) between the purification domain and the encoded polypeptide may be used to 
facilitate purification. One such expression vector provides for expression of a fusion 



BNSDOaD: <WO__019e38aA?J_? 



wo 01/96388 



PCTAJS01/18SS7 



50 

protein containing a polypeptide of interest and a nucleic acid encoding 6 histidine 
residues preceding a thioredoxin or an enterokinase cleavage site. The histidine residues 
facilitate purification on IMIAC (inunobilized metal ion afBnity chromatography) as 
described in Porath, J. et al. (1992, Prot Exp. Purif J263-281) while the enterokinase 

5 cleavage site provides a means for purifying the desired polypeptide fiom the fusion 
protein. A discussion of vectors which contain fusion proteins is provided in Kroll, D. J. 
et al, (1993; DNA Cell Biol. 72:441-453). 

In addition to recombinant production methods, polypeptides of the 
invention, and Augments thereof, may be produced by direct peptide synthesis usmg 

10 solid-phase techniques (Meirifield J. (1963) J, Am. Chem. Soc. 55:2149-2154). Protein 
synthesis may be performed using manual techniques or by automation. Automated 
synthesis may be achieved, for example, using Applied Biosystems 431A Peptide 
Synthesizer (Perkin Ehner). Alternatively, various fragments may be chemically 
synthesized separately and combined using chemical methods to produce the foil length 

15 molecule. 

Antibody Compositions, Fragments Thereof and Other BnwiNC Agents 

' Accoiding to another aspect, the present invention further provides 
binding agents, such as antibodies and antigen-binding fragments thereof, that exhibit 
immunological binding to a tumor polypeptide disclosed herem, or to a portion, variant 

20 or derivative thereof. An antibody, or antigen-binding fragment thereof is said to 
"specifically bind,'* "immunogically bind," and/or is "inununologically reactive" to a 
polypeptide of thte invention if it reacts at a detectable level (within, for example, an 
ELISA assay) with the polypeptide, and does not react detectably with unrelated 
polypeptides under similar conditions. 

25 Immunological binding, as used in this context, generally refers to the 

non-covalent interactions of the type which occur between an immunoglobulin 
molecule and an antigen for which the immunoglobulm is specific. The strength, or 
affinity of immunological binding interactions can be ejqpressed in terms of the 
dissociation constant (K<0 of the mteraction, wherein a smaller Kd represents a greater 

30 afBnity. Immunological bmdmg properties of selected polypeptides can be quantified 



BNSOOCID: -dWO P19638&A?.U» 



WO01/9&I88 



PCTAJSOl/18557 



51 

using methods well known in art. One such mefliod entails measuring the rates of 
antigen-binding site/antigen complex formation and dissociation, wherein those mtes 
depend on the concentrations of the complex partners, the afGnity of the interaction, 
and on geometric parameters that equally influence the rate in botfa directions. Thus, 

5 bofli the "on rate constant" (Kon) and the "off rate constant" (Kofr) can be determined by 
calculation of the concentrations and the actual rates of association and dissociation. 
The ratio of Koff /K<in enables cancellation of all parameters not related to affinity, and is 
thus equal to the dissociation constant K4. See, generally, Davies et al. (1990) Annual 
Rev. Biochem. 59:439-473. 

10 An "antigen-bindmg site," or "binding portion" of an antibody refers to 

the part of the immunoglobulin molecule that participates in antigen binding. The 
antigen binding ^te is formed by ammo acid residues of the N-terminal variable ("V") 
regions of the heavy ("H**) and light ("L") chains. Three highly divergent stretches 
within the V regions of the heavy and li^ chains are referred to as *1iypervariable 

15 regions" which are interposed between more conserved flanking stretches known as 
"fitimework regions," or "FRs"» Thus the t^m "FR" refers to amino add sequences 
which are naturally found between and adjacent to bypervariable regions in 
immunoglobulins. In an antibody molecule, the three bypervariable regions of a light 
chain and the three hypervariable regions of a heavy chain are disposed relative to each 

20 other in three dimensional space to form an antigen-binding surface. The antigen- 
binding surface is complementary to the three-dimensional surface of a bound antigen, 
and the three hypervariable regions of each of the heavy and light chains are referred to 
as "complementarity-determining regions," or "CDRs." 

Binding agents may be furth^ capable of diffs^tiating between 

25 patients with and without a cancer, such as colon cancer, using die representative assays 
provided herem. For example, antibodies or other binding agents that bmd to a tumor 
protem will preferably generate a signal indicating the presence of a cancer in at least 
about 20% of patients with the disease, more preferably at least about 30% of patients. 
Alternatively, or in addition, the antibody will generate a negative signal indicating the 

30 absence of the disease in at least about 90% of individuals without the cancer. To 
determine whether a binding agent satisfies this requirement, biological samples {e.g.. 



<WO__0196388Ai.L?' 



wo 01796388 



52 



PCTAJSOl/1^7 



blood, sera, sputum, urine and/or tumor biopsies) from patients with and without a 
cancer (as deteimined using standard clinical tests) may be assayed as described herein 
for the presence of polypeptides that Wnd to the binding a&aat. Preferably, a 
statistically significant numbw of samples with and without the disease will be assayed. 
5 Each binding agMil should satisfy the above criteria; however, Aose of ordinary skill in 
ibe art will recognize that Wnding agents may be used in combination to imi»ove 
sensitivity. 

Any agent that satisfies die above requirements may be a binding agent. 
For example, a binding agent may be a ribosome, with or without a pq)tide component, 
10 an RNA molecule or a polypeptide. In a preferred embodiment, a binding agent is an 
antibody or an antigen-binding fragment thereof. Antibodies may be prepared by any 
of a variety of techniques known to those of ordinary skill in the art. See, e.g., Hariow 
and Lane, Antibo<£es: A Laboratory Manual, Cold Spring Harbor Laboratory, 1988. In 
gCTieral, antibodies can be produced by cell culture tediniques, including the generation 
15 of monoclonal antibodies as described herein, or via transfection of antibody genes into 
suitable bacterial or mammalian cell hosts, in order to allow for the production of 
lecombinaitf antibodies. In one tedmique, an immunogaa comprising the polypeptide 
is mitially injected iiito any of a wide variety of mammals ie.g.. mice, rats, rabbits, 
sheep or goats). In this step, flie polypeptides of tiiis invention may serve as tiie 
20 immunogen without modification. Alternatively, particularly for relatively short 
polypeptides, a si^)erior immune response may be elicited if the polypeptide is joined to 
a carrier protein, such as bovine serum albumin or keyhole limpet hemocyanin. The 
immunogen is injected into the animal host, preferably according to a predetermined 
schedule incorporating one or more booster immunizations, and the animals are bled 
25 periodically. Polyclonal antibodies specific for die polypeptide may then be purified 
fit>m such antisera by, for example, affinity chromatography using the polypeptide 
coupled to a suitable solid support. 

Monoclonal antibodies specific for an antigenic polypeptide of interest 
may be prepared, for example, usnig die technique of KoUer and Milstein, Eur. J. 
30 Immunol d:511-519, 1976, and improvements thereto. Briefly, these mettiods involve 
the preparation of immortal cell lines cq)able of producing antibodies having tiie 



BNSOOCn: <WO___01ie388MU> 



wo 01796388 



53 



PCT/USOl/18557 



desired specificity (i.e., reactivity vdtfa &e polypeptide of interest). Such cell lines may 
be produced, for example, from spleen cells obtained from an animal immunized as 
described above. The spleen cells are then immortalized by, for example, fusion with a 
myeloma cell fusion partner, preferably one that is syngendc-jwitfa the immunized 

5 animal. A variety of fusion techniques may be employed. For example, the spleen cells 
and myeloma cells may be combined with a nonionic detergent for a few minutes and 
then plated at low density on a selective medium that supports the growth of hybrid 
cells, but not myeloma cells. A preferred selection technique uses HAT (hypoxanthine, 
aminopterin, thymidine) selection. After a sufficient time, usually about 1 to 2 weeks, 

10 colonies of hybrids are observed. Single colonies are selected and their culture 
sqpematants tested for binding activity against the polypeptide. Hyhridomas having 
high reactivity and specifidty are preferred. 

Monoclonal antibodies may be isolated fix)m the supematants of growing 
hybridoma colonies, hi addition, various techniques may be employed to enhance the 

1 5 yield, such as injection of the hybridoma cell line into the peritoneal cavity of a suitable 
vertebrate host, such as a mouse. Monoclonal antibodies may then be harvested fixmi 
the asdtes fluid or the blood. Contaminants may be removed fix>m fte antibodies by 
conventional techniques, such as chromatography, gel filtration, precipitation, and 
extraction. The polypeptides of tiiis invention may be used in the purification process 

20 in, for example, an affinity chromatography step. 

A number of therapeutically useful molecules are known in tiie art which 
comprise antigen-binding sites that are capable of exhibiting immunological binding 
properties, of an antibody molecule. The proteolytic enzyme papain preferentially 
cleaves IgG molecules to yield several fi:agments, two of which (the 'T(ab)" firagments) 

25 each comprise a covalent heterodimer that includes an intact antigen-binding site. The 
enzyme pepsin is able to cleave IgG molecules to provide several fiiagments, including 
the "F(ab')2 " fira^ent which comprises both antigen-Wnding sites. An "^Fv" fragment 
can be produced by preferential proteolytic cleavage of an IgM, and on rare occasions 
IgG or IgA inmiuno^obulin molecule. Fv fragments are, however, more commonly 

30 derived using recombinant techniques known in the art. The Fv fragment includes a 
non-covalent Vh::Vl heterodimer including an antigen-binding site which retains much 



BNSDOCm: <W0 0196388A?JL> 



wo 01/96388 



54 



PCT/USOl/18557 



of the antigen recognition and binding capabilities of the native antibody molecule. 
Inbar et al, (1972) Proc. Nat Acad. Sci. USA 69:2659-2662; Hochman et al. (1976) 
Biochem 15:2706-2710; and Efariich et al. (1980) Biochem 19:4091-4096. 

A single chain Fv ("sFv") polypeptide is a coTOlentJy linked Vh::Vl 

5 heterodimer which is ejqwcssed firom a gene fusdon including Vh- and Vt-encoding 
genes linked by a peptide-encoding linker. Huston et al. (1988) Proc. Nat. Acad. Sci. 
USA 85(16):5879-5883. A number of methods have been described to discern chemical 
structures for converting the naturally aggregated-but chemically separated-light and 
heavy polypeptide chains from an antibody V region into an sFv molecule which will 

10 fold into a three dimensional structure substantially similar to the structure of an 
antigen-binding site. See, e.g., U.S. Pat. Nos. 5,091,513 and 5,132,405, to Huston et al.; 
and U.S, Pat. No. 4,946,778, to Ladner et al. 

Each of the above-described molecules includes a heavy chain and a 
light chain CDR set, le^pectively interposed between a heavy chain and a light chain 

15 FR set i?^ch provide siqpport to Ihe CDRS and define the spatial relationship of the 
CDRs relative to each other. As used herein, the tenn "CDR set** refers to the three 
hypervariable regions of a heavy or ligiht chain V region. Proceeding fix)m the N- 
teiminus of a heavy or light chain, these regions are denoted as "CDRl," "CDR2,'' and 
"CDR3" respectively. An antigen-binding site, therefore, includes six CDRs, 

20 comprismg the CDR set from each of a heavy and a light chain V region. A polypeptide 
comprising a single CDR, (e.g., a CDRl, CDR2 or CDR3) is referred to herein as a 
"molecular recogm'tion unit." Crystallographic analysis of a number of antigen-antibody 
complexes has dfemonstrated that the amino acid residues of CDRs form extejosive 
contact with bound antigen, wherem the most extensive antigen contact is with the 

25 heavy chain CDRS. Thus, the molecular recognition units are primarily responsible for 
the spedficity of an antigen-bmding site. 

As used herem, the term "FR set" refers to the four flanking amino acid 
sequences which frame the CDRs of a CDR set of a heavy or light chain V region. 
Some FR residues may contact bound antigen; however, FRs are prfanarily responsible 

30 for folding the V region into flie antigen-bm^ng site, particularly the FR residues 
directly adjacent to the CDRS. Within FRs, certain amino residues and certain structural 



wo 01/%388 



PCT/US01/18S57 



55 

features are very highly conserved. Jn this regard, all V region sequences contain an 
internal disulfide loop of around 90 amino acid residues. When the V regions fold into a 
binding-site, the CDRs are displayed as projecting loop motifs which form an antigen- 
bmdmg sur&ce. It is gesierally recognized that there are conserved structural regions of 

5 FRs which influence the folded shape of the CDR loops into certain "canomcal" 
structures-regardless of the precise CDR amino acid sequence. Further, certain FR 
residues are known to participate in non-covalent interdomain contacts which stabilize 
the interaction of the antibody heavy and light chains. 

A number of "humanized" antibody molecules comprising an antigen- 

1 0 binding site derived from a non-human immunoglobulin have he&x described, including 
chimeric antibodies having rodent V regions and their associated GDRs fused to human 
constant domains (Winter et al. (1991) Nature 349293-299; Lobuglio et al. (1989) 
Proc. Nat. Acad. Sd. USA 86:4220-4224; Shaw et al. (1987) J Immunol. 138:4534- 
4538; and Brown et al. (1987) Cancer Res. 47:3577-3583), rodent CDRs grafted into a 

15 human supporting FR prior to fusion with an appropriate human antibody constant 
domain (Riechmann et al. (1988) Nature 332:323-327; Verhoeyen et al. (1988) Science 
239:1534-1536; and Jones et al. (1986) Nature 321:522-525), and rodent CDRs 
supported by recombinantly veneered rodent FRs (European Patent Publication No. 
519,596, published Dec. 23, 1992), These "humanized" molecules are designed to 

20 minimize unwanted inrniunological response toward rodent antihuman antibody 
molecules which limits the duration and effectiveness of therapeutic applications of 
those moieties m human recipients. 

As used herein, the terms "veneered FRs" and "recombinantly veneered 
FRs" refer to the selective replacement of FR residues from, e.g., a rodent heavy or light 

25 chain V region, with human FR residues in order to provide a xenogeneic molecule 
comprismg an antigen-binding site v^ich retains substantially all of the native FR 
polypeptide folding structure. Veneering techniques are based on the understanding that 
the ligand binding characteristics of an antigen-binding site are determined primarily by 
the structure and relative disposition of the heavy and light chain CDR sets wilfam &e 

30 antigen-binding surface. Davies et al. (1990) Ann. Rev. Biochem. 59:439-473. Thus, 
antigen binding specificity can be preserved in a humanized antibody only wherein the 



<W0 P106388MLL> 



wo 01/96388 



PCT/CS01/185S7 



56 

CDR structures, their interaction with each other, and their interaction with the rest of 
4e V region domains are carefully maintamed. By usuig veneering techniques, exterior 
(e.g., solvent-accessible) FR residues which are leadfly encountered by the immune 
system are selectively replaced with human residues to provide a hybrid molecule that 
5 comprises either a weakly unmunogenic, or substantially non-immunogenic veneered 
surface. 

The process of veneering makes use of the available sequence data for 
human antibody variable domains compiled by Kabat et al,, in Sequences of ftnotems of 
Immunological Interest, 4th ed., (U.S. Dept. of Health and Human Services, U,S. 

10 Government Printing Office, 1987), updates to flie Kabat database, and other accessible 
U.S. and foreign databases (both nucleic acid and protein). Solvent accessibilities of V 
region amino acids can be deduced from the known three-dimensional structure for 
human and murme antibody fragments. There are two general steps m veneering a 
murine antigen-bindmg site. Initially, the FRs of the variable domains of an antibody 

15 molecule of mterest are compared with correspondmg FR sequences of human variable 
domains obtamed from flie above-identified sources. The most homologous human V 
regions are then compared residue by residue to correspondmg murine ammo acids. Hie 
residues in the murine FR vAich differ from the human counterpart are replaced by the 
residues present in the human moiety using recombinant tedmiques well known m flie 

20 art. Residue switching is only carried out with moieties which are at least partially 
exposed (solvent accessible), and care is exercised in the replacement of amino add 
residues which may have a significant effect on the tertiary structure of V region 
domains, such as proline, glycine and charged amino acids. 

In this manner, the resultant "veneered" murine antigen-binding sites are 

25 thus designed to retain the murine CDR residues, the residues substantially adjacent to 
the CDRs, the residues identified as buried or mostly buried (solvent inaccessiblej, the 
residues believed to participate in non-covalent (e.g., electrostatic and hydrophobic) 
contacts between heavy and light chain domains, and the residues from conserved 
structural regions of the FRs which arc believed to influence the "canonical" tertiary 

30 structures of the CDR loops. These design criteria are then used to prepare recombmant 
nucleotide sequences which combine the CDRs of both the heavy and light cham of a 



BNSDOCID: 'eWO__0196388A?X> 



WO01/9(»388 



PCTA)S01/18557 



57 

murine antigen-binding site into human-appeaiing FRs that can be used to transfect 
mammalian cells for the expresaon of recombinant human antibodies vAAoh exhibit the 
antigen specificity of the murine antibody molecule. 

In another embodiment of the mvention, monoclonal antibodies of the 

5 presmt invention may be coiqpled to one or more therapeutic agents. Suitable agents in 
this regard include radionuclides, differentiation inducers, drugs, toxins, and derivatives 
thereof. Preferred radionuclides include '^I, '^I, "^e, ^^'^e, ^"At, and 
^'^Bi. Preferred drugs include methotrexate, and pyrimidine and purine analogs. 
Preferred differentiation inducers include phorbol esters and butyric acid. Preferred 

10 toxins include ricin, abrin, diptheria toxin, cholera toxin, gelonin, Pseudomonas 
exotoxin. Shigella toxin, and pokeweed antiviral protein. 

A tfaenqpeudc agent may be coiq)led («.g., covalently bonded) to a 
suitable monoclonal antibody either directly or indirectly (e.g., via a linker group). A 
direct reaction between an agent and an antibody is possible when each possesses a 

IS substituent ceipable of reacting with the other. For example, a nucleophilic group, such 
as an amino or sulfhydiyl group, on one may be capable of reacting with a carbonyl- 
containing group, such as an anhydride or an acid haUde, or with an alkyl group 
containmg a good leaving group a halide) on the ofter. 

Alternatively, it may be desirable to couple a thempeutic agent and an 

20 antibody via a linker group. A linker group can function as a spacer to distance an 
antibody from an agent in order to avoid interference with binding capabilities. A 
linker group can also serve to increase the chemical reactivity of a substituent on an 
agent or an antibody, and thus increase the coupling efficiency. An increase in 
chemical reactivity may also facilitate the use of agents, or functiona] groiq)s on agents, 

25 whidi otherwise would not be possible. 

It will be evident to those skilled in the art that a variety of bifunctional 
or polyfunctional reagents, both homo- and hetero-fimctioiud (such as those described 
in the catalog of the Pierce Oiemical Co., Rockford, IL), may be employed as the linker 
group. Coupling may be effected, for example, through amino groups, carboxyl groups, 

30 sulfliydryl groups or oxidized carbohydrate residues. There are numerous references 
describing such methodology, e.g,, U.S. Patent No. 4,671,958, to Rodwell et al. 



BNSOOCtCh <V«Q__019e3B8A?JL> 



wo 01/96388 



58 



PCTAJSOl/18557 



Where a therapeutic agent is more potent when free from the antibody 
portion of the immunoconjugates of &e present invention, it may be desirable to use a 
linker group which is cleavable during or upon internalization into a cell. A number of 
different cleavable linker groups have been described! The mechanisms for the 

S intracellular release of an agent from these linker groups include cleavage by reduction 
of a disulfide bond (e.g.y U,S, Patent No. 4,489,710, to Spitler), by irradiation of a 
photolabile bond U.S. Patent No. 4,625,014, to Senter et al.), by hydrolysis of 
derivatized amino acid side chains (e.g., U.S. Patent No. 4,638,045, to Kohn et al.), by 
serum complement-mediated hydrolysis (e.g., U.S. Patent No, 4,671,958, to Rodwell 

10 et al.), and acid-catalyzed hydrolysis U.S. Patent No. 4,569,789, to Blattier et al.). 

It may be desirable to couple more than one agent to an antibody. In one 
embodiment, multiple molecules of an agent are coupled to one antibody molecule. In 
another embodiment, moi^ than one type of agent may be coupled to one antibody. 
Regardless of flie particular embodiment, immunoconjugates vnOx more than one agent 

15 may be prepared in a variety of ways. For example, more than one agent may be 
coupled directly to an antibody molecule, or linkers that provide multiple sites for 
attadunentcanbeused. Alternatively, a carrier can be used. 

A carrier may bear the agents in a variety of ways, including covalent 
bonding either directly or via a linker group. Suitable carriers include proteins such as 

20 albumins (e.g., U.S. Patent No. 4,507,234, to Kato et al.), peptides and polysaccharides 
such as aminodextran (e.g., U.S. Patent No. 4,699,784, to Shih et al.). A carrier may 
also bear an agent by noncovalent bonding or by encapsulation, such as within a 
liposome vesicle (e.g., U.S. Patent Nos. 4,429,008 and 4,873,088). Carriers specific for 
radionuclide agents include radiohalogenated small molecules and chelating 

25 compounds* For example, U.S. Patent No. 4,735,792 discloses representative 
radiohalogenated small molecules and their synthesis. A radionuclide chelate may be 
formed from chelating compounds that inchide those containing nitrogen and sulfrir 
atoms as die donor atoms for binding the metal, or metal oxide, radionuclide. For 
ejcample, U.S. Patent No. 4,673,562, to Davison et al. discloses representative chelating 

30 compounds and tibeir ^thesis. 



BNSDOCID: <W0 .01863aaASLL> 



wo 01/96388 



59 



PCT/USOl/18557 



T Cell Compositions 

The present invention, in another aspect, provides T cells spedfic for a 
tumor polypeptide disclosed herein, or for a variant or derivative thereof. Such cells 
may generally be prepared in vitro or ex vivo^ using standard procedures. For example, 

S T cells may be isolated fiom bone marrow, peripheral blood,, or a fraction of bone 
manow or peripheral blood of a patient, using a commercially available cell separation 
system, such as the Isolex™ System, available from Nexell Therapeutics, Inc. (Irvine, 
OA; see also U.S. Patent No. 5,240,856; U.S. PatentNo. 5,215,926; WO 89/06280; WO 
91/16116 and WO 92/07243). Altematively, T cells may be derived from related or 

10 unrelated humans, non-human mammals, cell lines or cultures. 

T cells may be stimulated with a polypeptide, polynucleotide encoding a 
polypeptide and/or an antigen presenting cell (APC) that expresses such a polypeptide. 
Such stimulation is performed und^ conditions and for a time sufficient to permit the 
graeration of T cells that are specific for the polypeptide of interest. Preferably, a 

IS tumor polypeptide or polynucleotide of the invention is present within a delivery 
vehicle, such as a microsphere, to facilitate the generation of specific T cells. 

T cells are considered to be specific for a polypeptide of Ihe present 
invention if tiie T cells specifically proliferate, secrete cytokines or kill target cells 
coated witii the polypeptide or expressing a gene encoding the polypeptide. T cell 

20 specificity may be evaluated using any of a variety of standard techniques. For 
example, within a chromium release assay or proliferation assay, a stimulation index of 
more than two fold increase in lysis and/or proliferation, compared to negative controls, 
indicates T cell specificity. Such assays may be performed, for example, as described 
in Caien et al.. Cancer Res. 5-^:1065-1070, 1994. Altematively, detection of the 

25 im>liferation of T cells may be accomplished by a variety of known techniques. For 
example, T cell proliferation can be detected by measuring an increased rate of DNA 
synthesis (e.g., by pulse-labeling cultures of T cells with tritiated tiiymidine and 
measuring the amount of tritiated thymidine mcoiporated into DNA). Contact with a 
tumor polypeptide (100 ng/ml - 100 ^g/ml, preferably 200 ng/ml - 25 ^g/ml) for 3 - 7 

30 days will typically result m at least a two fold increase in proliferation of the T cells. 
Contact as described above for 2-3 hours should result in activation of the T cells, as 



BNSDOaO: <WO__019e388A?JU> 



wo 01/96388 



60 



PCTADSOl/18557 



measured using standard cytokine assays in which a two fold increase in the level of 
cytokine release (e.g„ TNF or DFN-y) is mdicative of T cell activation {see Coligan et 
al.. Current Protocols in Immunology, vol. 1, Wiley Intersdence (Greene 1998)). T 
ceDs that have been activated in response to a tumor polypeptide, polynucleotide or 

5 polypeptide-expressing APC may be CD4'' and/or CD8*. Tumor polypeptide-spedfic T 
cells may be expanded using standard techniques. Within preferred embodiments, the T 
cells are derived from a patient, a related donor or an unrelated donor, and are 
administered to the patient following stimulation and expansion. 

For therapeutic purposes, CD4"'" or T cells that proliferate in 

10 response to a tumor polypeptide, polynucleotide or APC can be expanded in number 
either in vitro or in vivo. Proliferation of such T cells in vitro may be accomplished m a 
variety of ways. For example, the T cells can be re-exposed to a tumor polypeptide, or 
a short peptide corresponding to an immunog^c portion of sudi a polypeptide, with or 
without the addition of Tcdl growfli factors, such as mterleukin-2, and/or stunulator 

15 cells that synthesize a tumor polypeptide. Alternatively, one or more T. cells that 
proliferate in the presence of the tumor polypeptide can be expanded in number by 
cloning. Methods for clonmg cells are well known in the art, and include limiting 
dilution. 

T Cell Receftor Compositions 

20 The T ceU receptor (TCR) consists of 2 diflferent, highly variable 

polypeptide chains, termed the T-cell receptor a and p chains, that are linked by a 
disulfide bond (Janeway, Travers, Walport. Jmmunohiology. Fourth Ed., 148-159. 
Elsevier Science Ltd/Garland Publishing. 1999). The ot/p heterodimer complexes with 
&e invariant CD3 chains at the cell membrane. This complex recognizes specific 

25 antigenic peptides bound to MHC molecules. The enormous divCTsity of TCR 
. specificities is generated much like immunoglobulin diversity, through somatic gene 
learrangement. The P cham genes contain over 50 variable (V), 2 diversity (D), over 10 
joming (J) segments, and 2 constant region segments (C). The a chain genes contam 
over 70 V segments, and over 60 J segments but no D segments, as well as one C 

30 segment. During T cell development in the thymus, the D to J gene rearrangement of 



BNSOOCID: <WO__019e388A^X> 



wo 01/96388 



61 



PCT/USOl/18557 



the p chain occurs^ followed by the V gene segment reanangemeot to the DJ. This 
functional VDJp exon is tamscribed and spliced to jom to a Cp. For the a chain, a Vo 
gene segment rearranges to a Ja gene segment to create the functional exon that is then 
transcribed and spliced to the Ca- Diversity is further mcreased during the 

5 recombination process by the random addition of P and N-nucleotides between the V, 
D, and J segments of the p cham and between the V and J segments in the a chain 
(Janeway, Travers, Walport. Imnmnohiology. Fourth Ed., 98 and 150. Elsevier Science 
Ltd/Garland Publishing. 1999). 

The present invention, m another aspect, provides TCRs specific for a 

10 polypeptide disclosed herein, or for a variant or derivative thereof. In accordance with 
the present invention, polynucleotide and amino acid sequences are provided for the V- 
J or V-D-J junctional regions or parts tiiereof for the alpha and beta chains of the T-cell 
receptor which recognize tumor polypeptides described herein. In general, this aspect 
of the invention relates to T-<:ell receptors which recognize or bind tumor polypeptides 

IS presented in the context of MHC. Jn a preferred embodiment the tumor antigens 
recognized by the T-cell receptors comprise a polypeptide of tiie present invention. For 
example, cDNA flooding a TCR specific for a colon tumor peptide can be isolated 
firom T cells specific for a tumor polypeptide using standard molecular biological and 
recombinant DNA techniques. 

20 This invention furtiier includes tiie T-cell receptors or analogs thereof 

having substantially the same function or activity as the T-cell receptors of this 
invention which recognize or bind tumor polypeptides. Such receptors include, but are 
not limited to, a fragment of tiie receptor, or a substitution, addition or deletion mutant 
of a T-icell receptor provided herem. This invention also encompasses polypeptides or 

25 peptides that are substantially homologous to the T-cell receptors provided herem or 
that retain substantially the same activity. The term "analog" includes any pfotem or 
polypeptide having an amino acid residue sequence substantially identical to the T-cell 
receptors provided herein in which one or more residues, preferably no more than 5 
residues, more preferably no more than 25 residues have been conservatively 

30 substituted with a functionally similar residue and which displays the functional aspects 
of the T-cell receptor as described herein. 



wo 01/96388 PCTAJSOl/18557 

62 

The present invention further provides for suitable mammalian host 
cells, for example, non-specific T cells, that are transfected with a polynucleotide 
encoding TCRs specific for a polypeptide described herein, thereby rendering the host 
cell specific for the polypeptide. The a and p chains of the TCR may be contained on 

5 separate expression vectors or alternatively, on a single expression vector that also 
contains an mtemal ribosome entry site (IRES) for c{q)-indep^dent translation of the 
gene downstream of the SUES. Said host cells expressing TCRs specific for the 
polypeptide may be used, for example, for adoptive immunotherapy of colon cancer as 
discussed further below. 

10 In further aspects of the present invention, cloned TCRs specific for a 

polypeptide recited herein m^ be used in a kit for the diagnosis of colon canc^. For 
example, the nucleic acid sequence or portions thereof, of colon tumor-specific TCRs 
can be used as probes or idmers for the detection of expression of the rearranged genes 
encoding the specific TCR in a biological sample. Thmfore, the present invention 

1 S further provides for an assay for detecting messenger RNA or DNA encoding the TCR 
specific for a polypeptide. 

ft 

Pharmaceutical Compositions 

In additional embodiments, the present invention concerns formulation 
of one or more of the polynucleotide, polypeptide, T-cell, TCR, and/or antibody 

20 compositions disclosed herein in pharmaceutically-acceptable carriers for 
administration to a cell or an animal, either alone, or in combination with one or more 
other modalities of therapy. 

It will be understood tiiat, if desired, a composition as disclosed herein 
may be administered in combination with other agents as well, sudi as, eg., other 

25 proteins or polypeptides or various pharmaceutically-active agents. In fact, there is 
virtually no limit to other components that may also be included, given that the 
additional agents do not cause a significant adverse effect upon contact with the target 
cells or host tissues. The compositions may thus be delivered along with various other 
agents as required in the particular instance. Such compositions may be purified jBx)m 

30 host cells or other biological sources, or alternatively may be chemically synthesized as 



BNSDOCID: <W0 ^0t9638eA^t> 



wo 01/96388 PCT/USOl/18557 

63 

described herein. Likewise, such compositions may further comprise substituted or 
derivatized KNA or DNA compositions. 

Therefore, in ano&er aspect of the present invention, pharmaceutical 
compositions are provided comprising one or more of the polynucleotide, polypeptide, 
5 antibody, TCR, and/or T-cell compositions described herein in combination with a 
physiologically acceptable carrier. In certain preferred embodiments, the 
pharmaceutical compositions of the invention comprise immunogenic polynucleotide 
and/or polypeptide compositions of the invention for use in prophylactic and tfaeraputic 
vacdne applications. Vaccine preparation is generally described in, for example, M.F. 

10 Powell and MJ. Newman, eds., "Vaccine Design (the subunit and adjuvant appioaohy* 
Plenum Press (NY, 1995). Generally^ such compositions will comprise one or more 
polynucleotide and/or polypeptide compositions of the present invention in combination 
with one or more immunostimulants. 

It will be apparent that any of the pharmaceutical compositions desoribed 

IS herem can contain pharmaceutically acceptable salts of the polynucleotides and 
polypeptides of the invention. Such salts can be prepared, for example, from 
pharmaceutically acceptable non-toxic bases, including organic bases (e.g., salts of 
primary, secondary and tertiary amines and basic amino acids) and inorganic bases 
(e.g., sodium, potassium, lithium, ammonium, calcium and magnesium salts). 

20 In another embodiment, illustrative immunogenic compositions, eg., 

vaccine compositions, of the present invention comprise DNA encoding one or more of 
the polypeptides as described above, such ibat the polypeptide is generated in situ. As 
noted ahoye, the polynucleotide may be administ^ted witfam any of a variety of delivery 
systems known to those of ordinary skill in the art Indeed, numerous gene delivery 

25 techniques are well known ia the art, such as those desocibed by RoUand, Crit Rev. 
TTiercp. Drug Carrier System i5:143-198, 1998, and refewmces cited therein. 
Appropriate polynucleotide eTqiresdon systems will, of course, contain the necessary 
regulatory DNA regulatory sequences for expression in a patient (such as a suitable 
promoter and terminating signal). Alternatively, bacterial delivery systems may involve 

30 the administration of a bacterium (such as Bacillus-Calmette-Guerrin) that expresses an 
immunogenic portion of the polypeptide on its cell surface or secretes such an epitope. 



BNSDOCID: <yVO__0198388A?JU> 



wo 01/96388 



64 



PCT/USOl/18557 



Therefore, in certain embodinients, polynucleotides encoding 
immunogenic polypeptides described herein are introduced into suitable mammalian 
host cells for expression using any of a number of known viral-based systems. In one 
illustrative embodiment, retroviruses provide a convenient andjeffective platform for 

5 gene delivery systems. A selected nucleotide sequence encoding a polypeptide of the 
present invention can be inserted into a vector and packaged in retroviral particles using 
techniques known in the art. The recombinant virus can then be isolated and delivered 
to a subject. A number of illustrative retroviral systems have been described (e.g., U.S. 
Pat. No. 5,219,740; Miller and Rosman (1989) BioTechniques 7:980-990; MUler, A. D. 

10 (1990) Human Gene Thempy 1:5-14; Scarpa et al. (1991) Virology 180:849-852; Bums 
et al. (1993) Proc. Natl. Acad. Sci. USA 90:8033-8037; and Boris-Lawie and Temin 
(1993) Cur. Opin. Genet Develop. 3:102-109. 

In addition, a number of illustradve adenovirus-based systems have also 
been described. Unlike retroviruses which mtegrate uito the host genome, adenoviruses 

15 persist extrachromosomally thus minimizing the risks associated wifli insertional 
mutagenesis (Haj-Ahmad and Graham (1986) J. Virol. 57:267-274; Bett et al. (1993) J. 
Virol. 67:5911-5921; Mittereder et al. (1994) Human Gene Thraapy 5:717-729; Sefli et 
al. (1994) J. Virol. 68:933-940; Bair et al. (1994) Gene Therapy 1 :51-58; Beikner, K. L. 
(1988) BioTechniques 6:616-629; and Rich et al. (1993) Human Gene Therapy 4:461- 

20 476). 

Various adeno-associated virus (AAV) vector systems have also been 
developed for polynucleotide delivery. AAV vectors can be readily constructed using 
techniques well known in the art. See, e.g., U.S. Pat. Nos. 5,173,414 and 5,139,941; 
International Publication Nos. WO 92/01070 and WO 93/03769; Lebkowski et al. 

25 (1988) Molec. Cell. Biol. 8:3988-3996; Vincent et al. (1990) Vaccines 90 (Cold Sprmg 
Harbor Laboratory Press); Carter, B. J. (1992) Current Opinion in Biotechnology 3:533- 
539; Muzyczka, N. (1992) Current Topics in Microbiol, and hnmunol. 158:97-129; 
Kotin, R. M. (1994) Human Gene Therapy 5:793-801; Shelling and Smith (1994) Gene 
Therapy 1 :165-169; and Zhou et al. (1994) J. Exp. Med. 179:1867-1875. 

30 Additional viral vectors useful for delivering the polynucleotides 

encoding polypeptides of the present invention by gene transfer include those derived 



BNSDOCID:<WO. 



1196388^^1^ 



PCT/OSOl/18557 



65 

fix>m the pox family of viruses, such as vaccinia virus and avian poxvirus. By way of 
example, vaccinia vims recombinants expressing the novel molecules can be 
constructed as follows. The DNA encoding a polypeptide is first inserted into an 
^propriate vector so that it is adjacent to a vaccinia promoter >»id flanking vacdnia 
5 DNA sequences, such as the sequence encoding thymidine kinase (TK). This vector is 
&en used to transfect cells vMch are simultaneously uifected with vaccinia. 
Homologous recombination serves to insert the vaccuiia promoter plus the gene 
^coding the polypeptide of interest into the viral genome. The resulting TK.sup.(*) 
recombinant can be selected by culturing the cells in the presence of S- 

10 bromodeoxyuridine and picking viral plaques resistant thereto. 

A vaccinia-based infection/transfection system can be conveniratly used 
to provide for inducible, transient e7q)ression or coexpression of one or more 
polypeptides described herein in host cells of an organism. In fliis particular system, 
cells are jSrst infected in vitro with a vaccinia virus recombinant that encodes fte 

15 bacteriophage T7 RNA polymerase. This polymerase displays exquisite specificity in 
that it only transcribes templates bearing T7 promoters. Following infection, cells are 
transfected with the polynucleotide or polynucleotides of interest, driven by a T7 
promoter. The polymerase expressed in the cytoplasm fi-om the vaccinia virus 
recombinant transcribes the transfected DNA into RNA which is then translated into 

20 polypeptide by the host translational machinery. The method provides for high level, 
tran^ent, cytoplasmic production of large quantities of RNA and its translation 
products. See, e.g., Ekoy-Stein and Moss, Ptoc. Natl. Acad. Sci. USA (1990) 87:6743- 
6747; Fuerst et al. Proc. Natl. Acad. Sci. USA (1986) 83:8122-8126. 

Alternatively, avipoxviruses, such as the fowlpox and canarypox viruses, 

25 can also be used to deliver the coding sequences of interest. Recombinant avipox 
viruses, expressing immunogens fiom mammalian pathogens, are known to confer 
protective immunity when administered to non-avian species. Hie use of an Avipox 
vector is particularly desirable in human and other inammalian species since memb^ 
of the Avipox genus can only productively replicate in susceptible avian species and 

30 therefore are not infective m mammalian cells. Methods for producing recombinant 
Avipoxviruses are known in the art and employ genetic recombination, as described 



wo 01/96388 



66 



PCT/US01/18S57 



above wifli respect to the production of vaccinia viruses. See, e.g., WO 91/12882; WO 

89/03429; and WO 92/03545. 

Any of a number of alphavirus vectors can also be used for delivery of 

polynucleotide compositions of the present invention, such as those vectors described in 
5 U.S. Patent Nos. 5,843,723; 6,015,686; 6,008,035 and 6,015,694. Certam vectors based 

on Venezuelan Equine Encephalitis (VEE) can also be used, illustrative examples of 

which can be found in U.S. Patent Nos. 5,505,947 and 5,643,576, 

Moreover, molecular conjugate vectors, such as the adenovirus chimeric 

vectors described in Michael et al. J. Biol. Chem. (1993) 268:6866-6869 and Wagner et 
10 al. Proc. Natl. Acad. Sci. USA (1992) 89:6099-61 03, can also be used for gene delivery 

under the invention. 

Additional illustrative information on these and other known viral-based 

delivery systems can be found, for example, in Fisher-Hoch et al., iVac. NatL Acad Sci. 

USA «d:317-321, 1989; Flexner et 3l.,Ann KY, Acad ScL 569:^6-103, 1989; Flexner 
15 et al.. Vaccine «:17-21, 1990; U.S. Patent Nos. 4,603,112. 4,769330, and 5,017,487; ' 

WO 89/01973; U.S. Patent No. 4,777,127; GB 2,200,651; EP 0,345,242; WO 91/02805; 

Berkner, Biotechniques (y:616-627, 1988; Rosenfeld et al.. Science 252:431-434, 1991; 

KoUs et al., Proc. NatL Acad Sci USA 97:215-219, 1994; Kass-Eisler et aL, Proc. Natl. 

Acad Sci USA P0:1 1498-1 1502, 1993; Guzman et al.. Circulation «*:2838-2848, 1993; 
20 and Guzman et al., Cir. Res, 73:1202-1207, 1993. 

In certain embodiments, a polynucleotide may be integrated mto the 

genome of a target cell. This integration may be in the specific location and orientation 

via homologous ifecombination (gene replacement) or it may be integrated in a random, 

non-specific location (gene augmentation). In yet fiulher embodiments, the 
25 polynucleotide may be stably maintained in the cell as a separate, episomal segment of 

DNA. Such polynucleotide segments or "episomes" encode sequences sufiBcient to 

permit maintenance and replication independent of or in synchronization with the host 

cell cycle. The manner m which the expression construct is delivered to a cell and 

where in tiie cell the polynucleotide remains is dependent on the type of expression 
30 construct employed. 



BNSDOCID: <W0 ^019638BAajj> 



wo 01/96388 



67 



PCTAJSOl/18557 



In another embodiment of the invention, a polynucleotide is 
administered/delivered as ""naked" DNA, for example as described in Ulmer et al.^ 
Science 2JP:1745-1749, 1993 and reviewed by Cohen, Science 2JP:1691-1692, 1993. 
The uptake of naked DNA may be mcreased by coating &e DNA onto biodegradable 

S beads, which are efficiently transported into the cells. 

In still another embodiment, a composition of the present invention can 
be delivered via a particle bombardment approach, many of which have been described. 
In one illustrative example, gas-driven particle acceleration can be achieved with 
devices such as those manufactured by Powderject Pharmaceuticals PLC (Oxford, UK) 

10 and Powdegect Vaccmes Inc. (Madison, WI), some examples of which are described in 
U.S. Patent Nos. 5,846,796; 6,010,478; 5,865,796; 5,584,807; and EP Patent No. 0500 
799. This sqpproach offers a needle-fiee delivery approach wherein a dry powder 
formulation of microscopic particles, such as polynucleotide or polypeptide particles, 
are accelerated to high speed within a helium gas jet generated by a hand held device, 

1 5 propelling Hit particles into a target tissue of interest. 

In a related embodim^t, other devices and methods that may be useful 
for gas-driven needle-less injection of compositions of the present invention include 
those provided by Bioject, Inc. (Portland, OR), some examples of which are described 
m U.S. Patent Nos. 4,790,824; 5,064,413; 5,312,335; 5,383,851; 5,399,163; 5,520,639 

20 and 5,993,412. 

According to another embodiment, the pharmaceutical compositions 
described herein will comprise one or more immunostimulants in addition to the 
immunogenic polynucleotide, polypqjtide, antibody, T-cell, TCR, and/or APC 
compositions of this invention. An inununostimulant refers to essentially any substance 

25 that enhances or potentiates an immune response (antibody and/or cell-mediated) to an 
exogenous antigen. One preferred type of immunostimulant comprises an adjuvant. 
Many adjuvants contain a substance designed to p-otect the antigen fix>m rapid 
catabolism, such as aluminum hydroxide or mineral oil, and a stunulator of immune 
responses, such as lipid A, Bortadella pertussis or Mycobacterium tuberculosis derived 

30 proteins. Certain adjuvants are commercially available as, for example, Freund's 
Incomplete Adjuvant and Complete Adjuvant (Difco Laboratories, Detroit, MI); Merck 



BNSDOCID: ^__0t963aaA3JL> 



wo 01^6388 



68 



PCT/US01/18S57 



Adjuvant 65 (Merck and Company, Inc., Rahway, NJ); AS-2 (SmithKline Beecham, 
Philadelphia, PA); alumimim salts such as aluminum hydroxide gel (alum) or aluminum 
phosphate; salts of calcium, iron or zinc; an insoluble suspension of acylated tyrosine; 
acylated sugars; cationically or anionically derivadzed polysacdiaiides; 

5 polypho^hazenes; biodegradable microspheres; monophosphqiyl lipid A and quil A. 
Cytokines, such as GM-CSF, interleukin-2, -7, -12, and other like growth fectors, may 
also be used as adjunrants. 

Within certain embodiments of the invention, the adjuvant composition 
is preferably one that induces an immune response predominantly of the Thl type. 

10 High levels of Thl-type cytokines (e.g., IFN^, TNFo, 11^2 and IL-12) tend to fevor the 
induction of cell mediated immune responses to an administered antigen. In contrast, 
high levels of Th2-type cytokines (e.g„ IL-4, IL-5, IL-6 and lL-lO) tend to favor the 
induction of humoral immune responses. Following application of a vaccine as 
provided herem, a patient will support an immune response that iiicludes Thl- and Th2- 

15 type responses. Withm a preferred embodiment, in which a response is predominantly 
Thl-type, fte level of Thl-type cytokines will increase to a greater extent than the level 
of Th2-type cytokines. The levels of these cytokmes may be readily assessed using 
standard assays. For a review of the families of cytokines, see M osmann and Cofi&nan, 
Ann, Rev, Immunol 7:145-173, 1989. 

20 Certain preferred adjuvants for eliciting a predominantly Thl-type 

response include, for example, a combination of monophosphoryl lipid A, preferably 3- 
de-O-acylated monophosphoryl lipid A, together with an aluminum salt. MPL* 
adjuvants are available from Corixa Corporation (Seattle, WA; see, for example, US 
Patent Nos. 4,436,727; 4,877,611; 4,866,034 and 4,912,094). CpG-containing 

25 oligonucleotides (in which the CpG dinucleotide is unmethylated) also induce a 
piedominantly Thl response. Such oligonucleotides are well known and are described, 
for example, in WO 96/02555, WO 99/33488 and U.S. Patent Nos. 6,008,200 and 
5,856,462. Inununostimulatory DNA sequences are also described, for example, by 
Sato et al.. Science 273:252, 1996. Another preferred adjuvant comprises a saponin, 

30 such as Quil A, or derivatives thereof, including QS21 and QS7 (Aquila 
Biophaimaceuticals Inc., Framingham, MA); Escin; Digitonin; or Gypsophila or 



BNSDOCID: <W0 ^0Y9638aASLU> 



wo 01/9^88 



PCT/USOl/18557 



69 

Chenopodium qtdnoa saponins . Other piefeired formulations include more than one 
saponin in the adjuvant combinations of the present invration, for example 
- combinations of at least two of tfie following group comprising QS21, QS7, Quil A» 
esdn, or digHonin. 

5 Alternatively the saponin formulations may be combined with vaccine 

vehicles composed of chitosan or other polycationic polymers, polylactide and 
polylactide-co-glycolide particles, poly-N-acetyl glucosamine-based polymer matrix, 
particles composed of polysaccharides or chemically modified polysaccharides, 
liposomes and lipid-based particles, particles composed of glycerol monoesters, etc. 

10 The saponins may also be formulated in the presence of cholesterol to form particulate 
structures such as liposomes or ISCOMs. Furthermore, fte saponins may be formulated 
together with a polyoxyethylene ether or ester, in either a non-particulate solution or 
suspension, or in a particulate structmre such as a paucilamelar liposome or ISCOM. 
The saponins may also be fonnulated with excipients such as Carbopol^ to increase 

15 viscosity, or may be formulated in a dry powder form with a powder excipient such as 
lactose. 

In one preferred embodimrat, the adjuvant system includes die 
combination of a monophosphoryl lipid A and a s^nin derivative, such as the 
combination of QS21 and 3D-MPL® adjuvant, as described in WO 94/00153, or a less 
20 reactogenic composition where the QS21 is quenched with cholesterol, as described in 
WO 96/33739. Other preferred formulations comprise an oil-in-water emulsion and 
tocopherol. Another particxJarly preferred adjuvant formulation employing QS21, 3D- 
MPL adjuvant and tocopherol in an oil-in-water emulsion is described in WO 
95/17210. 

25 Another enhanced adjuvant system involves the combination of a CpG- 

containing oligonucleotide and a saponin derivative particulariy tiie combination of 
CpO and QS21 is disclosed m WO 00/09159. Preferably the formulation additionaUy 
comprises an oil in water emulsion and tocopherol. 

Additional illustrative adjuvants for use m the pharmaceutical 

30 compositions of the invention include Montanide ISA 720 (Seppic, France), SAF 
(Chiron, California, United States), ISCOMS (CSL), MF-59 (Chiron), the SBAS series 



BNSDOCID: <WO ^0196388^«J^ 



wo 01/96388 



PCTAJSOl/18557 



70 

of adjuvants SBAS-2 or SBAS-4, avaUable jfrom SmitbKline Beecham, Rixensart, 
Belgium), Detox (Enhanzyn®) (Corixa, Hanulton, Ml), RC-529 (Corixa, Hamilton, 
MT) and other aminoalkyl glucosaminide 4-phosphates (AGPs), such as those 
described in pending U.S. Patent Application Serial Nos. 08/853,826 and 09/074,720, 
5 the disclosuies of which are incorporated herein by reference in fheir entireties, and 
polyoxyethylene ether adjuvants such as those described in WO 99/52549A1 . 

Other preferred adjuvants include adjuvant molecules of the general 

formula 

(I): HO(CH2CH20)„-A-R, 

10 wherein, n is 1-50, A is a bond or -C(0>, R is C1.50 alkyl or Phenyl C1.50 alkyl. 

One embodiment of the present invention consists of a vaccine 
formulation comprising a polyoxyethylene ether of general formula (I), wherein n is 
between 1 and 50, preferably 4-24, most preferably 9; the R component is Ci-50, 
preferably C4-C20 alkyl and most preferably C12 alkyl, and i4 is a bond. The 

15 concentration of tihe polyoxyethylene ethers should be in the range 0,1-20%, preferably 
from 0.1-10%, and most preferably in the range 0.1-1%. Preferred polyoxyethylene 
ethers are selected from flie following group: polyoxyethylene-9-lauryl ether, 
polyoxyethylene-9-steoryl ether, polyoxyefliylene-8-steoryl e&er, polyoxyethylene-4- 
lauryl ether, polyoxyethylene-354auryl ether, and polyoxyefliylene-23-lauryl ether. 

20 Polyoxyethylene ethers such as polyoxyethylene lauryl eftier are described in tiie Merck 
index (12* edition: entry 7717). These adjuvant molecules are described in WO 
99/52549. 

Thfe polyoxyethylene ether according to the general formula (I) above 
may, if desired, be combined with another adjuvant. For example, a preferred adjuvant 
25 combination is preferably with CpG as described m the pending UK patent application 
GB 9820956.2. 

According to another embodiment of this invention, an immimogenic 
composition desc^ed herein is delivered to a host via antigen presenting cells (APCs), 
such as dendritic cdls, macrophages, B cells, monocytes and other cells that may be 
30 engineered to be efficient APCs. Such cells may, but need not, be genetically modified 
to increase the cspm^ for presenting the antigen, to improve activation and/or 



BNSDOaO: <W0 .0196386ASLL> 



wo 01/96388 



71 



PCT/USOl/18557 



maintenance of the T cell response, to have anti-tumor effects per se and/or to be 
inmmnologically compatible ^vith the receiver {ie., matched HLA haplotype). APCs 
may generally be isolated fiom any of a variety of biological fluids and organs, 
including tumor and peritumoral tissues, and may be autologous^^ogeneic, syngeneic 
5 or xenogeneic cells. 

Certam preferred embodimrats of the present invention use dendritic 
cells or progenitors thereof as antigen-presenting cells. Dendritic cells are highly potent 
APCs (Banchereau and Steinman, Nature 3P2:245-251, 1998) and have been shown to 
be effective as a physiological adjuvant for eliciting prophylactic or therapeutic 

10 antitumor immunity (see Timmerman and Levy, Amt Rev. Med 50:507-529, 1999). h 
general, dendritic cells may be identified based on their typical shape (stellate in situ, 
with marked cytoplasmic processes (dendrites) visible in vitroX their ability to take up, 
process and present antigens with high ef35ciency and their ability to activate nafve T 
cell responses. Dendritic cells may, of course, be engineered to express specific cell- 

15 surface receptors or ligands that are not commonly found on dendritic cells in vivo or ex 
vivo, and such modified dendritic cells are contemplated by the present invention. As 
an alternative to dendritic cells, secreted vesicles antigen-loaded dendritic cells (called 
exosomes) may be used within a vaccine (see Zitvogel et al.. Nature Med 4:594-600, 
1998). 

20 Dendritic cells and progenitors may be obtained firom peripheral blood, 

bone maiTOW, tumor-infiltrating cells, pmtumoral tissues-infiltrating cells, lymph 
nodes, spleen, skin, umbilical cord blood or any other suitable tissue or fluid. For 
example, -dendritic cells may be differentiated ex vivo by adding a combination of 
cytokmes such as GM-CSF, IL-4, IL-13 and/or TNFa to cultures of monocytes 

25 harvested from peripheral blood. Alternatively, CD34 positive cells harvested firom 
peripheral blood, umbilical cord blood or bone marrow may be differentiated into 
dendritic cells by adding to the culture medium combinations of GM-CSF, IL-3, TNFa, 
CD40 ligand, LPS, flt3 ligand and/or other compound(s) that induce differentiation, 
maturation and proliferation of dendritic cells. 

30 Dendritic cells are conveniently categorized as "inunature" and "mature" 

cells, which allows a simple way to discriminate between two well characterized 



BNSOOaO: ^01963BaA^_L> 



wo 01/96388 



72 



PCTAJSOl/18557 



phenotypes. However, this nomenclature should not be construed to exclude all 
possible intermediate stages of differentiation. Immature dendritic cells are 
characterized as APC with a high capacity for antigen uptake and processing, which 
conelates with the high expression of Fey receptor and mannose receptor. The mature 
5 phenotype is typically characterized by a lower expression of these markers, but a high 
expression of cell surface molecules responsible for T cell activation such as class I aiwi 
class n MHC, adhesion molecules (e.g-, CD54 and CDl 1) and costimulatory molecules 
(e.g., CD40, CD80, CD86 and 4-lBB). 

APCs may generally be transfected with a polynucleotide of the 
10 invention (or portion or other variant thereof) such that the encoded polypeptide, or an 
immunogenic portion thereof, is expressed on the cell surface. Such transfection may 
take place ex vivo, and a pharmaceutical composition comprising such transfected cells 
may then be used for therapeutic purposes, as described herein. Alternatively, a gene 
delivery vehicle that targets a dendritic or other antigen presenting cell may be 
15 administered to a patient, resulting in transfection that occurs in vivo. In vivo and ex 
vivo transfection of dendritic cells, for example, may generally be performed using any 
methods known in fte art, such as those described in WO 97/24447, or the gene gun 
approach described by Mahvi et al.. Immunology and cell Biology 75:456-460, 1997. 
Antigen loading of dendritic cells may be achieved by mcubating d«idritic cells or 
20 progenitor ceUs with the tumor polypeptide, DNA (naked or within a plasmid vector) or 
RNA; or with antigen-expressing recombinant bacterium or viruses (e.g., vaccinia, 
fowlpox, adenovirus or lentivirus vectors). Prior to loading, the polypeptide may be 
covalently conjugated to an immunological partner that provides T cell help {e,g,, a 
carrier molecule). Alternatively, a dendritic cell may be pulsed with a non-conjugated 
25 immunological partner, separately or in the presence of the polypeptide. 

While any suitable carrier known to those of ordinary skill in the art may 
be employed in the pharmaceutical compositions of this invention, the type of carrier 
will typically vary depending on the mode of administration. Compositions of the 
present invention njay be formulated for any appropriate manner of administration, 
30 including for example, topical, oral, nasal, mucosal, intravenous, intracranial, 
intraperitoneal, subcutaneous and intramuscular administration. 



BNSDOCIDc <VW) ^0196388A2J.> 



wo 01/96388 



PCTAJSOl/18557 



73 

Carriers for use Avilfain such phannaceutical compositions are 
biocompatible, and may also be biodegradable. In certain embodiments, the 
fonnuladon preferably provides a relatively constant level of active component release. 
In odier embodiments, however, a more rapid rate of release immedi£itely upon 
5 administration may be desired. The fomiulation of such compositions is well wiflun the 
level of ordinary skill in the art using known techniques. Illustrative carriers useful in 
this regard include microparticles of poly(lactide-co-glycolide), polyacrylate, latex, 
starch, cellulose, dextran and the like. Other illustrative delayed-release carriers 
include supramolecular biovectors, which comprise a non-liquid hydrophilic core (e.g., 

10 a cross-linked polysaccharide or oligosaccharide) and, optionally, an extonal layer 
compriang an amphiphilic compound, such as a phospholipid {see e.g., U.S. Patent No. 
5,151,254 and PCX appHcations WO 94/20078, WO/94/23701 and WO 96/06638). The 
amount of active compound contained within a sustained release formulation depends 
upon the site of implantation, the rate and expected duration of release and tiie nature of 

15 the condition to be treated or prevented. 

In another illustrative embodiment, biodegradable microspheres (e.g., 
polylactate polyglycolate) are employed as carriers for the compositions of this 
invention. Suitable biodegradable microspheres are disclosed, for example, in U.S. 
Patent Nos. 4,897,268; 5,075,109; 5,928,647; 5,811,128; 5,820,883; 5,853,763; 

20 5,814,344, 5,407,609 and 5,942,252. Modified hepatitis B core protein carrier systems, 
such as described in WO/99 40934, and references cited therein, will also be useful for 
many applications. Another illustrative carrier/delivery system employs a carrier 
comprising particulate-protein complexes, such as those described in U.S. Patent No. 
5,928,647, which are capable of inducing a class I-restricted cytotoxic T lymphocyte 

25 responses in a host 

In another illustrative embodiment, calcium phosphate core particles are 
employed as carriers, vaccine adjuvants, or as controlled release matrices for the 
compositions of this invention. Exemplary calcium phosphate particles are disclosed, 
for example, in published patent application No. WO/0046147. 

30 The phannaceutical compositions of the invention wifl often further 

comprise one or more buffers (e,g., neutral buffered saline or phosphate buffered 



wo 01/^388 



74 



PCT/US01/18SS7 



saline^ carbohydrates (e.g., glucose, mannose, sucrose or dextrans), maimitol, proteins, 
polypq)tides or amino acids such as glycine, antioxidants, bacteriostats, chelating 
agents sudi as EDTA or glutathione, adjuvants alunrinum hydroxide), solutes that 
render the formulation isotonic, hypotonic or weakly hypertonic vdlh Ihe blood of a 

5 recipient, suspending agents, thickening agents and/or preservatives. Alternatively, 
compositions of Ihe present invention may be formulated as a lyophilizate. 

The pharmaceutical compositions described herdn may be presented in 
unit-dose or multi-dose containers, such as sealed ampoules or vials. Such containers 
are typically sealed in such a way to preserve the sterility and stability of the 

10 formulation until use. In general, formulations may be stored as suspensions, solutions 
or emulsions in oily or aqueous vehicles. Alternatively, a phannaceutical composition 
may be stored in a freeze-dried condition requiring only the addition of a sterile liquid 
carrier immediately prior to use. 

The development of suitable dosing and treatment regimens for using the 

15 particular compositions described herein in a variety of treatment regimais, including 
e.g., oral, parenteral, intravenous, intranasal, and intramuscular administration and 
formulation, is well known in tiie art, some of which are briefly discussed below for 
g^eral purposes of illustration. 

In certain applications, the phannaceutical compositions disclosed herein 

20 may be delivered via oral administration to an animal. As such, these compositions 
may be foraiulated witii an inert diluent or with an assimilable edible carrier, or they 
may be enclosed in hard- or soft-shell gelatin capsule, or they may be compressed into 
tablets, or they may be incorporated directly with the food of the diet. 

The active compounds may even be incorporated with excipients and 

25 used in tiie form of ingestible tablets, buccal tables, troches, capsules, elixirs, 
suspensions, syrups, wafers, and the like (see, for example, Mathiowitz et aL, Nature 
1997 Mar 27;386(6623):410-4; Hwang et al, Crit Rev Ther Drug Camer Syst 
l998;15(3):243-84; U. S. Patent 5,641,515; U. S. Patent 5,580,579 and U. S. Patent 
5,792,451). Tablets, troches, pills, capsules and the like may also contam any of a 

30 variety of additional components, for example, a binder, such as gum tragacanth, 
acacia, cornstarch, or gelatin; excipients, such as dicalcium phosphate; a disintegrating 



BNSDOCID: ^WO ^01S6388A2LL> 



wo 01/96388 



75 



PCT/DSOl/18557 



agent, such as com staich, potato starch, algmic add and the like; a lubricant, such as 
magnesium stearate; and a sweetening agent, such as sucrose, lactose or saccharin may 
be added or a flavoring agent, such as pepp^mint, oil of vdntergreen, or cherry 
flavoring. When ib& dosage unit form is a capsule, it may contain, in addition to 
S materials of the above type, a liquid carrier. Various other materials may be present as 
coatings or to otherwise modify the physical form of the dosage unit. For instance, 
tablets, pills, or capsules may be coated with shellac, sugar, or both. Of course, any 
material used in preparing any dosage unit form should be pharmaceutically pme and 
substantially non-toxic in the amounts employed. In addition, the active compounds 

10 may be incorporated into sustained-release preparation and formiilations. 

Typically, these formulations will contain at least about 0.1% of the 
active compound or more, although the percentage of the active ingredient(s) may, of 
course, be varied and may conveniently be between about 1 or 2% and about 60% or 
70% or more of the weight or volume of the total formulation. Naturally, the amount of 

IS active compound(s) in each ther^eutically useful composition niay be prqpared is such 
a way that a suitable dosage will be obtained in any given unit dose of tiie compound 
Factors such as solubility, bioavailability, biological half-life, route of administration, 
product shelf life, as well as otiier pharmacological considerations will be contemplated 
by one skilled in tiie art of preparing such pharmaceutical formulations, and as such, a 

20 variety of dosages and treatment regimens may be desirable. 

For oral administration the compositions of the present invention may 
alternatively be incorporated with one or more excipients in the form of a mouthwash, 
dentifiice, buccal tablet, oral spray, or sublingual orally-administered formxilation. 
Alternatively, the active ingredient may be incorporated into an oral solution such as 

25 one containing sodium borate, glycerin and potassium bicarbonate, or dispersed in a 
dentifiice, or added in a therapeutically-efifective amount to a composition tiiat may 
include water, binders, abrasives, flavoring agents, foaming agents, and bumectants. 
Alternatively tiie compositions may be fashioned into a tablet or solution form tiiat may 
be placed under the tongue or otherwise dissolved in 'tiie mouth. 

30 In certain circumstances it will be desirable to deliver the pharmaceutical 

compositions disclosed herein parenterally, intravenously, intramuscularly, or even 



wo 01^88 PCTAJSOyiWS? 

76 

intr^eritoneally. Such approaches are well known to the skilled artisan, some of which 
are fhrflier described, for example, m U. S. Patent 5,543,158; U. S. Patent 5,641,515 
and U. S. Patent 5,399,363. hi certain embodiments, solutions of the active compounds 
as free base or pharmacologically acceptable salts may be prepared in water suitably 

5 mked with a surfactant, such as hydroxypropylcellulose. Mspersions may also be 
prepared in glycerol, liquid polyethylene glycols, and mixtures thereof and in oils. 
Under ordinary conditions of storage and use, these preparations generally will contain 
a preservative to prevent the growth of microorganisms. 

Illustrative pharmaceutical forms suitable for injectable use include 

10 sterile aqueous solutions or dispersions and sterile powders for the extemporaneous 
preparation of sterile injectable solutions or dispersions (for example, see U. S. Patent 
5,466,468). In all cases the form must be sterile and must be fluid to the extent that 
easy syringability exists. It must be stable under the conditions of manufacture and 
storage and must be preserved against the contaminating action of microorgani^s, 

15 sudi as bacteria and fungi. The carrier can be a solvent or dispersion medium 
contaming, for example, water, ethanol, polyol glycerol, propylene glycol, and 
liquid polyethylene glycol, and the like), suitable mixtures thereof, and/or vegetable 
oils. Proper flmdity may be maintamed, for example, by flie use of a coating, such as 
lecithin, by the maintenance of the required particle size in the case of dispersion and/or 

20 by the use of surfactants. The prevention of the action of microorganisms can be 
facilitated by various antibacterial and antifungal agents, for example, parabens, 
chlorobutanol, phenol, sorbic acid, thimerosal, and the like. In many cases, it will be 
preferable to include isotonic agents, for example, sugars or sodium chloride. 
Prolonged absorption of the injectable compositions can be brought about by the use in 

25 the compositions of agents delaymg absorption, for example, aluminum monostearate 
and gelatin. 

In one embodiment, for parenteral administration in an aqueous solution, 
the solution should be suitably buffered if necessary and the liquid diluent first rendered 
isotonic v^ith sufficient saline or ^ucose. These particular aqueous solutions are 
30 especially suitable for intravenous, intramuscular, subcutaneous and intraperitoneal 
administration. In this coimection, a sterile aqueous medium that can be employed vnH 



BNSDOCIO: <WO__010638aA3!JL;* 



wo 01/96388 



77 



PCTAJSOl/18557 



be known to those of sldll in flie art in light of Ae present disclosure. For example, one 
dosage may be dissolved in 1 ml of isotonic NaCl solution and either added to 1000 ml 
of hypodemoclysis fluid or injected at the proposed site of infusion, (see for example, 
"Remington's Phannaceutical Sciences" 15th Edition, pages 1035-1038 and 1570- 

5 1580). Some variation in dosage will necessarily occur depending on the condition of 
the subject being treated. Moreover, for human admmistration, preparations will of 
course preferably meet sterility, pyrogenicity, and the general safety and purity 
standards as required by FDA Office of Biologies standards. 

In another embodiment of the invention, the compositions disclosed 

10 herein may be fonnulaled in a neutral or salt form. niustmtive 
pharmaceutically-acceptable salts include the acid addition sahs (formed with the free 
amino groups of the protein) and which are formed wilfa inorganic adds such as^ for 
example, hydrochloric or phosphoric adds, or such organic adds as acetic, oxalic, 
tartaric, mandelic, and the like. Salts formed with the free carix)xyl groups can also be 

15 derived from inorganic bases such as, for example, sodium, potassium, ammonium, 
calcium, or ferric hydroxides, and such organic bases as isopropylamine, 
trimethylamine, histidine, procaine and the like. Upon formulation, sohitions wll be 
administered in a manner compatible with the dosage formulation and in such amount 
as is ^erapeutically efifective. 

20 The carriers can ftirther comprise any and all solvents, dispersion media, 

vehicles, coatings, diluents, antibacterial and antifimgal agents, isotonic and absorption 
delaying agents, buffers, carrier solutions, suspensions, colloids, and the like. The use 
of such media and agents for phannaceuticai active substances is wdl known m the art. 
Except insofar as any conventional media or agent is incompatible with the active 

25 ingredient, its use in ihe therapeutic conq>ositions is contemplated. Supplementary 
active ingredients can also be incorporated into the compositions. The phrase 
"pharmaceutically-acceptable" refers to molecular entities and compositions that do not 
produce an allergic or similar untoward reaction when administered to a human. 

In certain embodiments, the pharmaceutical compositions may be 

30 delivered by intranasal sprays, inhalation, and/or other aerosol delivery vdiicles. 
Methods for delivering genes, nucldc adds, and peptide compositions directiy to the 



BNSDOCID: <WCL__019e38BA;L.L> 



wo 01/96388 



78 



PCTAJSOl/18557 



lungs via nasal aerosol sprays has been described, e.g,, in U. S. Patent 5,7563^3 and U. 
S. Patent 5,804,212. Likewise, tbe deliv^ of drugs using intranasal microparticle 
lesins (Takenaga et ah, J Controlled Release 1998 Mar 2;52(l-2):81-7) and 
lysophosphatidyl-glycerol compounds (U. S. Patent 5,725,871) are also well-known in 

S the pharmaceutical arts. Likewise, illustrative transmucosal drug delivery in the form 
of apolytetrafluoroefheylene support matrix is described in U. S. Patent 5,780,045. 

In certain embodiments^ liposomes, nanocapsules, microparticles, lipid 
particles, vesicles, and the like, are used for the introduction of the compositions of the 
present invention into suitable host cells/oi^anisms. In particular, the compositions of 

10 the present invention may be formulated for delivery either encapsulated in a lipid 
particle, a liposome, a vesicle, a nanosphere, or a nanoparticle or the like. 
Alt^atively, compositions of the present invention can be bound, either covalently or 
non-covalently, to the sur&ce of such carrier vehicles. 

The fomiation and use of liposome and liposome-like preparations as 

15 potential drug carriers is generally known to those of skill in the art (see for example, 
Lasic, Trends Biotedmol 1998 Jul;16(7):307-21; Takakura, Nippon Rinsho 1998 
Mar,56(3):691-5; Chandran et al, Indian J Exp Biol. 1997 Aug;3S(8):801-9; MargaUt, 
Crit Rev Ther Drug Carrier Syst. 1995;12(2-3):233-61; U.S. Patent 5,567,434; U.S. 
Patent 5,552,157; U.S. Patent 5,565,213; U.S. Patent 5,738,868 and U.S. Patent 

20 5,795,587, each specifically incorporated herein by reference in its enturety). 

Liposomes have been used successfully with a number of cell types that 
are normally difficult to transfect by other procedures, including T cell suspensions, 
primary hepatocyte cultures and PC 12 cells (Renneisen et al, J Biol Chem. 1990 Sep 
25;265(27):16337-42; Muller et al, DNA Cell Biol. 1990 Apr,9(3):221-9). In addition, 

25 liposomes are free of tfie DNA length constraints that are typical of viral-based delivery 
systems. Liposomes have been used efTectively to introduce genes, various drugs, 
radiofherapeutic agents, enzymes, viruses, transcription factors, allosteric efiectors and 
the like, into a variety of cultured cell lines and animals. Furthermore, he use of 
liposomes does not appear to be associated wiA autoimmune responses or unaccqptable 

30 toxicity after systemic delivery. 



BNSDOQD: <WO .01963a8A^I_^ 



wo 01/96388 



PCT/USOl/18557 



79 

In certain embodiments, liposomes are formed fiom phospholipids that 
are dispersed in an aqueous medium and spontaneously form multilamellar concentric 
bilayer vesicles (also termed multilamellar vesicles (MLVs). 

Ahemadvely, in ofher embodiments, the invention provides for 
5 pharmaceutically-acceptable nanocapsule formulations of the compositions of the 
present invention. Nanocapsules can generally entrap compounds in a stable and 
reproducible way (see, for example, Quintanar-Guerrero et ah. Drug Dev Ind Pharm. 
1998 Dec;24(12):l 113-28). To avoid side effects due to intmcellular polymeric 
overloading, such ultrafine particles (sized around 0.1 jmi) may be designed using 
10 polymers able to be degraded in vivo. Such particles can be made as described, for 
example, by Couvreur eial., Crit Rev Ther Drug Carrier Syst 1988;5(l):l-20; zur 
Muhlen et a/., Eur J Pharm Biopham. 1998 Mai^45(2): 149-55; Zambaux et aL J 
Controlled Release. 1998 Jan 2;50(1-3):31-40; and U. S, Patent 5,145,684. 

Cancer Thkrapeutic Methods 

15 Immunologic approaches to cancer dierapy are based on the recognition 

that cancer cells can often evade the body's defenses against aberrant or foreign cells 
and molecules, and that these defenses might be therapeutically stimulated to regain the 
lost ground, e.g. pgs. 623-648 in Klein, Immunology (WileyJnterscience, New York, 
1982). Numerous recent observations that various immune effectors can directly or 

20 indirectly inhibit growth of tumors has led to renewed interest in this approach to cancer 
ther^y, e.g. Jager, et al.. Oncology 2001;60(l):l-7; Renner, et al., Ann Hematol 2000 
Dec;79(13):65lT9/ 

Four-basic cell types whose function has been associated with antitumor 
cell immunity and the elimination of tumor cells firom the body are: i) B-lymphocytes 

25 which secrete immunoglobulins into the blood plasma for identifying and labeling the 
nonself invader cells; ii) monocytes which secrete the complement proteins that are 
responsible for lysing and processing the immunoglobulin-coated target invader cells; 
iii) natural killer lymphocytes having two mechanisms for the destruction of tumor 
cells, antibody-dependent cellular cytotoxicity and natural killing; and iv) T- 

30 lymphocytes possessing antigen-specific receptors and having the capacity to recognize 



wo 01/96388 



PCT/US01/185S7 



80 

a tumor cell canying complementaiy marker molecules (Sehrdber, H., 1989, in 
Fundamental Immunology (ed). W. E. Paul, pp. 923-955). 

Cancer immunothen^y generally focuses on inducing humoral immune 
responses, cellular immune responses, or both. Moreover, it is_well established that 
5 induction of CD4^ T helper cells is necessary in ordar to secondarily mdnce either 
antibodies or cytotoxic CD8* T cells. Polypeptide antigais that are selective or ideally 
specific for cancer cells, particularly colon cancer cells, ofifer a powerful approach for 
inducing immune responses against colon cancer, and are an important aspect of the 
present invention. 

1 0 Therefore, in further aspects of the present invention, the pharmaceutical 

compositions described herein may be used to stimulate an immune response against 
cancer, particulaily for the immunother^y of colon cancer. Within such methods, the 
phannaceutical compositions described herein are admmistered to a patient, typically a 
warm-blooded animal, prefi^ably a human. A patient may or may not be afQicted with 

15 cancer. Pharmaceutical compositions and vaccines nmy be adnunistered eifli^ prior to 
or followmg surgical removal of primary tumors and/or treatment such as 
administration of radiotherq>y or conventional chetnoth^apeutic drugs. As discussed 
above, administration of the pharmaceutical compositions may be by any suitable 
method, including administration by intravenous, intraperitoneal, intramuscular, 

20 subcutaneous, intranasal, intradermal, anal, vaginal, topical and oral routes. 

Within certain embodiments, immunotherapy may be active 
immunotherapy, in which treatment relies on the in vivo stimulation of the endogenous 
host immune system to react against tumors with the administration of immune 
response-modifying agents (such as polypeptides and polynucleotides as provided 

25 herein). 

Withm other embodiments, immunotherapy may be passive 
immunottierapy, in vdiich treatment involves the delivery of agents wifli established 
tumor-immime reactivity (such as effector ceUs or antibodies) that can directly or 
indirectly mediate antitumor effects and does not necessarily depend on an intact host 
30 immune system. Examples of effector cells include T cells as discussed above, T 
lymphocytes (such as CDS* cytotoxic T lymphocytes and CD4^ T-helper tumor- 



BNSDOOD: <WO__019C388A?_I_> 



.- - ,.-^;*^^tt5*v.v«»a^•»l«o«»3^B^Ksa^•.■■•. 



WO 01/96388 PCT/US01/18SS7 

81 

infiltrating lymphocytes), killer cells (such as Natural Killer cells and lymphokine- 
activated killer cells), B cells and antigen-presenting cells (such as dendritic cells and 
macrophages) expressing a polypeptide provided herein. T cell receptors and antibody 
receptors specific for the polypeptides recited herein may be doned, expressed and 
5 transferred into other vectors or effector cells for adoptive immunotherapy. The 
polypeptides provided herein may also be used to generate antibodies or anti-idiotypic 
antibodies (as described above and m U.S. Patent No. 4,918,164) for passive 
immunotherapy. 

Monoclonal antibodies may be labeled with any of a variety of labels for 

10 desired selective usages in detection, diagnostic assays or tiia^peutic applications (as 
described in U.S. Patent Nos. 6,090365; 6,015,542; 5,843^98; 5.595,721; and 
4,708,930, hereby incorporated by reference in their entirety as if each was incorporated 
individually). In each case, the binding of tiie labelled monoclonal antibody to Ihe 
determinant site of the antigen will signal detection or delivery of a particular 

15 tiierapeutic agent to the antigenic determinant on the non-normal cell. A further object 
of tius invention is to provide the specific monoclonal antibody suitably labelled for 
achieving such desired selective usages thereof. 

Effector cells may generally be obtained in sufficient quantities for 
adoptive unmunotherapy by growth in vitro, as described herein. Culture conditions for 

20 expanding single antigen-specific effector cells to several billion in number with 
retention of antigen recognition in vivo are well known in the art. Such in vitro culture 
conditions typically use intermittent stimulation with antigen, often in the presence of 
cytokines, (such as IL-2) and non-dividing feeder cells. As noted above, 
immunoreactive polypeptides as provided herein may be used to rapidly expand 

25 antigen-specific T cell cultures in order to generate a sufficient number of cells for 
immunother^y. In particular, antigen-presenting cells, such as dendritic, macrophage, 
monocyte, fibroblast and/or B cells, may be pulsed with nnmunoreactive polypeptides 
or transfected with one or more polynucleotides using standard techniques well known 
m the art. For example, antigen-presenting cells can be transfected with a 

30 polynucleotide having a promoter appropriate for increasing expression m a 
recombinant vims or other expression system. Cultured effector cells for use in therapy 



BNSOOCID: <W0 ^01063B8A^Xp' 



wo 01^6388 



82 



PCTAJSQl/18557 



must be able to grow and distribute widely, and to survive long term in vivo. Studies 
have shown that cultured effector cells can be induced to grow in vivo and to survive 
long term in substantial numbers by repeated stimulation with antigen supplemented 
with IL-2 (see, for example, Oieev^ et al.. Immunological Reviews 757:177, 1997). 

5 Alternatively, a vector expressing a polypq)tide recited herein may be 

introduced into antigen presenting ceDs takra firom a patimt and clonally propagated ex 
vivo for transplant back into the same patient Transfected cells may be reintroduced 
into the patient using any means known in the art, preferably in sterile form by 
intravenous, intracavitary, intraperitoneal or intratumor administration. 

10 Routes and frequency of administration of the therapeutic compositions 

described herein, as well as dosage, will vary from individual to individual, and may be 
readily established using standard techniques. In general, the pharmaceutical 
compositions and vaccines may be administered by injection intracutaneous, 
intramuscular, intravenous or subcutaneous), intranasally (e.g., by aspiration) or orally, 

15 Preferably, between 1 and 10 doses may be administered over a 52 week period. 
Preferably, 6 doses are administered, at intervals of 1 month, and boosts vaccinations 
may be given periodically thereafter. Alternate protocols may be appropriate for 
individual patients. A suitable dose is an amount of a compound that, vAiea 
administered as described above, is capable of promoting an anti-tumor immune 

20 response, and is at least 10-50% above the basal (i.e., untreated) level. Such reqponse 
can be monitored by measuring the anti-tumor antibodies in a patient or by vaccine- 
dependent generation of cytolytic efiFector cells capable of killing the patient's tumor 
cells in vitro. SuCh vaccines should also be capable of causing an immune response that 
leads to an improved clinical outcome (e.g., more frequent remissions, complete or 

25 partial or longer disease-fiee survival) in vaccinated patients as compared to non- 
vaccinated patients. In general, for pharmaceutical compositions and vaccines 
comprising one or more polypeptides, the amount of each polypeptide present in a dose 
ranges from about 25 \ig to 5 mg per kg of host Suitable dose sizes will vary witfi flie 
size of the patient, but will typically range from about 0.1 mL to about 5 mL. 

30 In general, an appropriate dosage and treatment regunen provides the 

active compound(s) in an amount sufiBcient to provide then^utic and/or prophylactic 



BNSOOaD: «^WOL__O106388A2JL> 



wo 01/96388 



83 



PCT/USOl/18557 



benefit. Such a response can be monitored by establishing an improved clinical 
outcome (e.g., more fiequent remissions, complete or partial, txt longer disease-free 
survival) in treated patients as compared to non-treated patients, hu^eases in 
preexisting immime responses to a tumor protein generally comdate with an improved 
5 clinical outcome. Such immune responses may generally be evaluated usmg standard 
proliferation, cytotoxicity or cytokine assays, which may be performed using samples 
obtained from a patient before and after treatment 

Cancer Detection and Diagnostic Compositions, Methods and Kits 

In gen^, a cancer may be detected in a patient based on the presence 

10 of one or more colon tumor proteins and/or polynucleotides encoding such proteins in a 
biological sample (for example, blood, sera, sputum urine and/or tumor biopsies) 
obtained fix>m the patient In other words, such protems may be used as markers to 
indicate the presence or absence of a cancer such as colon cancer. In addition, such 
proteins may be useful for the detection of other cancers. The binding agents provided 

15 herein generally permit detection of the level of antigen that binds to the agent in the ^ 
biological sample. 

Polynucleotide primers and probes may be used to detect tiie level of 
n[iRNA encoding a tumor protein, which is also indicative of the presence or absence of 
a cancer. In general, a tumor sequence shoxild be present at a level that is at least two- 

20 fold, preferably three-fold, and more preferably fivefold or hi^er m tumor tissue than 
in normal tissue of the same type from which the tumor arose. E3q)ression levels of a 
particulac tumor sequence in tissue types different from that in which the tumor arose 
are irrelevant in certain diagnostic embodiments since the presence of tumor cells can 
be confirmed by observation of predetemained differential exp:ession levds, e.g^ 2- 

25 fold, S-fold, etc, in tumor tissue to expression levels in normal tissue of the same type. 

Other differential expression patterns can be utilized advantageously for 
diagnostic purposes. For example, in one aspect of the invention, overexpression of a 
tumor sequence in tumor tissue and normal tissue of the same type, but not in other 
normal tissue types, e.g, PBMCs, can be exploited diagnostically. In this case, the 

30 presence of metastatic tumor cells, for example in a sample taken from the circulation 



eNSD0CiD:<WO. 



wo 01^6388 



PCTAJS01/18SS7 



84 

or some other tissue site dififeieait from that in which the tumor arose, can be identified 
and/or confirmed by detecting expression of the tumor sequence in the sample, for 
example using RT-PCR analysis. In many instances, it Mvill be desired to enrich for 
tumor cells in the sample of int^est, e.g., PBMCs, using ceU_capture or other like 
5 techniques. 

There are a variety of assay fonnats known to those of ordinary skill in 
the art for using a binding agent to detect polypeptide markers in a sample. See, e.g., 
Hariow and Lane, Antibodies: A Laboratory Manual, Cold Spring Harbor Laboratory, 
1 988- In general, the presence or absence of a cancer in a patient may be deteimined by 

10 (a) contacting a biological sample obtained fiom a patient witfi a binding agent; (b) 
detecting in the sample a level of polypeptide that binds to flie bindixig agent; and (c) 
comparing the level of polypeptide with a predetermined cut-ofiT value. 

In a preferred embodiment, the assay involves the use of binding agent 
immobilized on a solid support to bind to and remove the polypeptide from the 

IS remainder of the sample. The bound polypeptide may then be detected using a 
detection reagent that contains a reporter group and specifically binds to the binding 
agent/polypeptide complex. Such detection reagents may comprise, for example, a 
binding agent that specifically binds to the polypeptide or an antibody or other agrat 
that specifically binds to the binding agent, such as an anti-immunoglobulin, protein G, 

20 protein A or a lectin. Altanatively, a competitive assay may be utilized, in which a 
polypeptide is labeled with a reporter gro\q> and allowed to bind to the nnmobilized 
, binding agent after incubation of the binding agent with the sample. The extent to 
which components of the sample inhibit the binding of the labeled polypeptide to the 
binding agent is indicative of the reactivity of fhe sample with the immobilized binding 

25 agent. Suitable polypeptides for use within such assays include full length colon tumor 
proteins and polypeptide portions thereof to which the binding agent binds, as described 
above. 

The solid support may be any material known to fliose of ordinary skill 
in the art to which the tumor protein may be attached. For example, the solid support 
30 may be a test well in a microtiter plate or a nitrocellulose or other suitable membrane. 
Alternatively, tiie support may be a bead or disc, such as glass, fiberglass, latex or a 



BNSDOCID: .rtVO__01fl638aA?.L? 



wo 01/96388 



PCT/USOl/18557 



85 

plastic material such as polystyrene or polyvinylcWoride. The support may also be a 
magnetic particle or a jRber optic sensor, such as fliose disclosed, for example, in U.S. 
Patent No. 5,359,681. The bmding agent may be immobilized on the solid support 
using a variety of techniques known to those of skill in the ^ which are amply 

5 described in the patent and scientific literature. In the context of the present invention, 
the term "immobilization" refers to both noncovalent association, such as adsorption, 
and covalent attachment (which may be a direct Imkage between the agent and 
functional groups on the support or may be a linkage by way of a cross-linking agent). 
Immobilization by adsorption to a well in a microtiter plate or to a membrane is 

10 preferred. In such cases, adsorption may be achieved by contacting the bindmg agent, 
in a suitable buffer, with flie solid suppOTt for a suitable amoimt of time. The contact 
tune varies with t^perature, but is typically between about 1 hour and about 1 day. In 
genera], contacting a well of a plastic microtiter plate (sudi as polystyrene or 
polyvinylchloride) with an amount of bindmg agent ranging fix)m about 10 ng to about 

15 10 ^ig, and preferably about 100 ng to about 1 (ig, is sufiBcient to immobilize an 
adequate amount of binding agent 

Covalent attachment of bindmg agsnt to a solid support may generally 
be achieved by first reactmg the support with a biiunctional reagent that will react with 
both the support and a functional group, such as a hydroxyl or amino group, on the 

20 binding agent. For example, the binding agent may be covalently attached to supports 
having an appropriate polymer coating using benzoquinone or by condensation of an 
aldehyde group on the support with an amine and an active hydrogen on the bindmg 
partner Xsee, e.g., Kerce Immunotechnology Catalog and Handbook, 1991, at 
A12-A13). 

25 In certain embodiments, the assay is a two-antibody sandwich assay. 

This assay may be performed by first contacting an antibody that has been inunobilized 
on a solid support, conunonly flie well of a naicrotiter plate, with the sample, such that 
polypeptides witfiin the sample are allowed to bind to the immobilized antibody. 
Unbound sample is then removed fiom the unmobilized polypeptide-antibody 

30 complexes and a detection reagent (preferably a second antibody capable of binding to 
a different site on the polypeptide) containing a reporter group is added. The amount of 



BMSOOOD: <WO__0196a88AeJL> 



WO01/9d388 



PCT/US01/185S7 



S6 

detection reagent that remains bound to the solid support is then determined using a 
method appropriate for the specific reporter group. 

More specifically, once the antibody is immobilized on the support as 
described above, the remaining protein binding sites on the si]^rt are typically 

5 blocked. Any suitable blocking agent known to those of ordinary skill in the art, such 
as bovine serum albumin or Tween 20™ (Sigma Chemical Co., St Louis, MO). The 
unmobilized antibody is then incubated with the sample, and polypeptide is allowed to 
bind to the antibody. The sample may be diluted with a suitable diluent, such as 
phosphate-buffered saline (PBS) prior to incubation. In general, an appropriate contact 

10 time (i.e., incubation time) is a period of time that is sufficient to detect the presence of 
polypeptide within a sample obtained from an individual with colon cancer at least 
about 95% of that achieved at equilibrium between bound and unbound polypeptide. 
Those of ordinary skill in the art will recognize that the time necessary to achieve 
equilibrium may be readily determined by assaying the level of binding that occurs over 

IS a period of time. At room temperature, an incubation time of about 30 minutes is 
generally sufficient 

Unboynd sample may then be ronoved by washing the solid support 
with an appropriate buffer, such as PBS containmg 0.1% Tween 20™. The second 
antibody, which contains a reporter group, may then be added to ttie solid support. 

20 Preferred reporter groups include those groups recited above. 

The detection reagent is then incubated with the immobilized antibody- 
polypeptide complex for an amount of time sufficient to detect the bound polypeptide. 
An appropriate ataiount of time may generally be determined by assaying the level of 
binding that occurs over a period of time. Unbound detection reagent is then removed 

25 and bound detection reagent is detected using the reporter group. The method 
^ployed , for detecting the reporter group depends upon the natore of the reporter 
gpmp. For radioactive groups, scintillation counting or autoradiogr^hic methods are 
generally appropriate. Spectroscopic methods may be used to detect dyes, luminescent 
groups and fluorescent groups. Biotin may be detected using avidin, coupled to a 

30 different rq)orter group (commonly a radioactive or fluorescent group or an enzyme). 
Enzyme reporter groups may generally be detected by the addition of substrate 



BNSOOCID: <WO___0196a88A«JL> 



wo 01/96388 



87 



PCT/OSOl/18557 



(generally for a specific period of time), followed by spectioscopic or other analysis of 
the reaction products. 

To detennine the presence or absence of a cancer, such as colon cancer, 
&e signal detected from the reporter group that remains bound to the solid support is 

5 generaUy compared to a signal that conresponds to a predetemMUed cut-off value. In 
one preferred embodiment, the cxrt-off value for the detection of a cancer is the average 
mean signal obtained when the immobilized antibody is incubated with samples fix)m 
patients without the cancer. In general, a sample generating a signal that is three 
standard deviations above the predetermined cut-off value is considered positive for the 

10 cancer. In an alternate preferred embodiment, the cnt-^ff value is determined using a 
Receiver Operator Curve, according to the method of Sackett et al.. Clinical 
Epidemiology: A Basic Science for Clinical Medicine, Utfle Brown and Co., 1985, 
p. 106-7. Brirfly, in this embodiment, the cut-off value may be detranined ftom a plot 
of pairs of true positive rates (/.e., sensitivity) and filse positive rates (100%- 

1 5 specificity) that correspond to each possible cut-off vahie for fte diagnostic test result 
The cut-oflF value on the plot that is the closest to the upper left-hand comer (i.e., the 
value that encloses the largest area) is the most accurate cut-off value, and a sample 
generating a signal that is higher than the cut-ofF value determined by this method may 
be considered positive. Alternatively, the cut-ofF value may be shifted to the left along 

20 the plot, to minimize the false positive rate, or to the right, to minimize the false 
negative rate. In general, a sample generating a signal ftat is higher than the cut-off 
value determined by this method is considered positive for a cancer. 

In a related embodiment, the assay is performed in a flow-through or 
strip test format, wherein the binding agent is inranobilized on a membrane, such as 

25 nitrocellulose. In die flow-through test, polypeptides within the sample bind to the 
immobilized binding agent as the sample passes through the membrane. A second, 
labeled bindmtg agent then binds to the binding agent-polypeptide complex as a solution 
containing the second binding agent flows through the membrane. The detection of 
bound second binding agent may then be performed as described above. In the strip test 

30 format, one end of the membrane to which binding agent is bound is immersed in a 
solution containing the sample. The sample migrates along the membrane through a 



BNSOOCIO: <WO__01Be388WLl^ 



WOOl/96388 PCT/US01/185S7 

88 

re^on containing second binding agent and to the area of immobilized binding agent 
Concentration of second binding agent at flie area of immobDized antibody indicates the 
presence of a cancer. Typically, the concratration of second binding agent at that site 
generates a pattern, such as a line, that can be read visually, llie absence of such a 

5 pattern indicates a negative result. In general, the amount of binding agent hnmobilized 
on the membrane is selected to generate a visually discernible pattern when the 
biological sample contains a level of polypeptide that would be sufficient to generate a 
positive signal in the two-antibody sandwich assay, in flie fonnat discussed above. 
Preferred binding agents for use in such assays are antibodies and antigen-binding 

10 fragments thereof- Preferably, the amount of antibody immobilized on the membrane 
ranges from about 25 ng to about l\ig, and more preferably from about 50 ng to about 
500 ng. Sudi tests can typically be peifoimed with a very sinall amount of biological 
sample. 

Of course, numerous other assay protocols exist that are suitable for use 

15 with the tumor proteins or binding agents of the present invention. The above 
descriptions are intended to be exemplary only. For example, it will be apparent to 
those of ordinary skill in the art that the above protocols may be readily modified to use 
tumor polypeptides to detect antibodies that bind to such polypeptides m a biological 
sample. The detection of such tumor protdn specific antibodies may correlate wifli the 

20 presence of a cancer. 

A cancer may also, or alternatively, be detected based on the presence of 
T cells that specifically react with a tumor protein m a biological sample. Within 
certain inethods, a biological sample comprising CD4* and/or CDS^ T cells isolated 
fix>m a patient is incubated with a tumor polypeptide, a polynucleotide encoding such a 

25 polypeptide and/or an AFC that expresses at least an immunogenic portion of such a 
polypeptide, and the presence or absence of specific activation of the T cells is detected. 
Suitable biological samples include, but are not limited to, isolated T cells. For 
example, T cells may be isolated firom a patient by routine techniques (such as by 
FicoU/Hypaque density gradient centrifiigation of peripheral blood lymphocytes). T 

30 cells may be incubated in vitro for 2-9 days (typically 4 days) at 37X with polypeptide 
(e.g., 5 - 25 ng/ml). It may be desirable to incubate another aliquot of a T ceU sample 



BNSOOaO: <W0 .0196388A^L^ 



wo 01/96388 



FCT/USOl/18557 



89 

in the absence of tumor polypeptide to serve as a control. For CD4* T cells, activation 
is preferably detected by evaluating proliferation of the TceUs, For 008"^ T cells, 
activation is preferably detected by evaluating cytolytic activity. A level of 
proliferation that is at least two fold greater and/or a level of cytolytic activity Chat is at 
5 least 20% greater than in disease-j&ee patients indicates the presence of a cancer in the 
patient 

As noted above, a cancer may also, or alternatively, be detected based on 
the level of mKNA encoding a tumor protein in a biological sample. For example, at 
least two oligonucleotide primers may be employed in a polymerase chain reaction 
10 (PGR) based assay to amplify a portion of a tumor cDNA derived from a biological 
sample, wherein at least one of the oligonucleotide primers is specific for (i.e., 
hybridizes to) a polynucleotide encoding the tumor protein. The amplified cDNA is 
dien separated and detected using techniques well known in tiie art, such as gel 
electrophoresis. 

15 Shnilarly, oligonucleotide . probes ftat specifically hybridize to a 

polynucleotide encoding a tumor protein may be used in a hybridization assay to detect 
the presence of polynucleotide encoding the tumor protein in a biological sample. 

To permit hybridization under assay conditions, oligonucleotide primers 
and probes should comprise an oligonucleotide sequence that has at least about 60%, 

20 preferably at least about 75% and more preferably at least about 90%, identity to a 
portion of a polynucleotide encoding a tumor protein of the invention that is at least 10 
nucleotides, and preferably at least 20 nucleotides, in length. Preferably, 
oligonucleotide primers and/or probes hybridize to a polynucleotide encoding a 
polypeptide described herein undex moderately stringent conditions, as defined above. 

25 Oligonucleotide primers and/or probes which may be usefiiUy employed m ttie 
diagnostic methods described herein preferably are at least 10-40 nucleotides in length. 
In a preferred embodiment, the oligonucleotide primera comjHise at least 10 contiguous 
nucleotides, more preferably at least 15 contiguous nucleotides, of a DNA molecule 
having a sequence as disclosed herein. Techniques for both PGR based assays and 

30 hybridization assays are well known in the art (see, for example, MuUis et al., Cold 



.019e3a8A^JL^ 



wo 01/96388 PCTAJSOl/18557 

90 

Spring Harbor Symp. Quant. Biol, 51:263, 1987; Erlich ed., FCR Technology. Stockton 
Press, NY, 1989). 

One preferred assay employs RT-PCR, in which PGR is applied in 
conjunction with reverse transcription. Typically, RNA is extrs^ed fiom a biological 

5 sample, such as biopsy tissue, and is reverse transcribed to produce cDNA molecules. 
PGR amplification using at least one specific primer generates a cDNA molecule, 
which may be separated and visualized using, for example, gel electrophoresis. 
Amplification may be performed on biological samples taken firom a test patient and 
fi-om an individual who is not afflicted with a cancer. The amplification reaction may 

10 be peifonned on several dilutions of cDNA spanning two orders of magnitude. A two- 
Md or greater increase in expression in several dilutions of the test patient sample as 
compared to the same dilutions of the non-cancerous sample is typically considered 
positive. . 

In another aspect of the present invention, cell capture technologies may 
15 be used in conjunction, with, for sample, real-time PGR to provide a more sensitive 
tool for detection of metastatic cells expressing colon tumor antigens. Detection of 
colon cancer cells in biological samples, e.g., bone marrow samples, peripheral blood, 
and small needle aspiration samples is desirable for diagnosis and prognosis in colon 
cancer patients. 

20 Inmiunomagnetic beads coated with specific monoclonal antibodies to 

surface cell mark^, or tetrameric antibody complexes, may be used to first enrich or 
positively select cancer cells in a sample. Various commercially available kits may be 
used, including Dynabeads® Epithelial Enrich (Dynal Biotech, Oslo, Norway), 
StemSep™ (StemCell Technologies, Inc., Vancouver, BC), and RosetteSep (StemCell 

25 Technologies). A skilled artisan will recognize that other methodologies and kits may 
also be used to enrich or positively select desired cell populations. Dynabeads® 
Epithelial Enrich contains magnetic beads coated with mAbs specific for two 
glycoprotein membrane antigens expressed on normal and neoplastic epithelial tissues. 
The coated beads may be added to a sample and the sample then applied to a magnet, 

30 tiiereby capturing the cells bound to the beads. The unwanted cells are washed away 
and the magnetically isolated cells eluted &om the beads and used m finlher analyses. 



BNSDOOD: <WO___01963MA2J^ 



wo 01/96388 



91 



PCTA)S01/185S7 



RosetteSep can be used to enrich cells directly fix>m a blood sample and 
consists of a cocktail of tetrameric antibodies that targets a variety of unvoted cells 
and crosslinks them to glycophorin A on red blood cells (RBC) present in the sample, 
fonning rosettes. When centriftiged over Ficoll, targeted cells pdlet along with the free 
5 RBC. The combination of antibodies in the depletion cocktail-determines which cells 
will be removed and consequently which cells will be recovered. Antibodies that are 
available include, but are not limited to: CD2, CDS, CD4, CDS, CDS, CDIO, CDllb, 
CD14, CD15, CD16, CD19, CD20, CD24, CD25, CD29, CD33, CD34, CD36, CD38, 
CD41, CD45, CD45RA, CD45RO, CD56, CD66B, CD66e, HLA-DR, IgE, and 
10 TCRop. 

Additionally, it is contemplated in Hxe present invention that mAbs 
specific for colon tumor antigens can be generated and used in a similar manner. For 
example, mAbs that bind to tumor-specific cell sur&ce antigens may be conjugated to 
magnetic beads, or formulated in a tetrameric antibody complex, and used to enrich or 

15 positively select metastatic colon tumor cells fi?om a sample. Once a sample is enriched 
or positively selected, ceUs may be lysed and KNA isolated. RKA may then be 
'subjected to RT-PCR analysis using colon tumor-specific primers in a real-time PGR 
assay as described herein. One skilled in the art will recognize that enriched or selected 
populations of cells may be analyzed by other methods (e,g. in situ hybridization or 

20 flow cytomctry). 

In another embodiment, the compositions described herein may be \ised 
as markers for the progression of cancer. In this embodiment, assays as described 
above for .the diagnosis of a cancer may be performed over time, and iht change in the 
level of reactive polypeptide(s) or polynucleotide(s) evaluated. For example, the assays 

25 may be performed every 24-72 hours for a period of 6 months to 1 year, and thereafter 
perfomied as needed. In gen^, a cancer is progressing in those patients in whom the 
level of polypeptide or polynucleotide detected increases over time. In contrast, the 
cancer is not progressing when the level of reactive polypeptide or polynucleotide either 
remains constant or decreases with time. 

30 Certain in vivo diagnostic assays may be performed directly on a tumor. 

One such assay involves contacting tumor cells with a binding agent. The bound 



BNSDOCID: <WO__oi9e38aA2LL> 



wo owesm 



PCTrtJS01/18K7 



92 

binding agent may then be detected direcdy or indirectly via a reporter groiq). Such 
binding agrats may also be used in histological applications. Alternatively, 
polynucleotide probes may be used within such applications. 

As noted above, to improve sensitivity, multiple tumor protein markers 

5 may be assayed within a given sample. It will be apparent that bindmg agents specific 
for different proteins provided herein may be combined within a single assay. Further, 
multiple primers or probes may be used concurrently. The selection of tumor protein 
markers may be based on routine experiments to determine combinations that results in 
optimal sensitivity. In addition, or alternatively, assays for tumor proteins provided 

10 herein may be combined with assays for other known tumor antigens. 

The present invention further provides kits for use within any of the 
above diagnostic methods. Such kits typically comprise two or more components 
necessary for performing a diagnostic assay. Components may be compounds, 
reagents, containers and/or equipment. For example, one container within a kit may 

15 contain a monoclonal antibody or firagm^t tiiereof that specifically binds to a tumor 
protein. Such antibodies or fi^gments may be provided attached to a siqpport material, 
as described above. One or more additional containers may enclose elements, sudi as 
reagents or buffers, to be used in the assay. Such kits may also, or altematively, contain 
a detection reagent as described above that contains a reporter group suitable for direct 

20 or indirect detection of antibody binding. 

Altematively, a kit may be designed to detect the level of mRNA 
encoding a tumor protein in a biological sample. Such kits generally comprise at least 
one oligonucleotide probe or primer, as described above, that hybridizes to a 
polynucleotide encoding a tumor protein. Such an oligonucleotide may be used, for 

25 example, within a PGR or hybridization assay. Additional components that may be 
present within such kits include a second oligonucleotide and/or a diagnostic reagent or 
container to facilitate the detection of a polynucleotide encoding a tumor pioteiiL 

The following Examples are offered by way of ilhistration and not by 
way of limitation. 

30 



BNSDOCID: <WO__Ol86388ASLL> 



wo 01/96388 



PCT/USOl/18557 



93 

EXAMPLES 
EXAMPLE 1 

IDENTIHCATION OF COLON TUMOR PROTEIN CDNAs 

5 This Example illustrates the identification of cDNA molecules encoding 

colon tumor proteins using PCR-based cDNA subtraction methodology. 

A modification of the Clontech (Palo Alto, CA) PCR-Select™ cDNA 
subtraction methodology vms employed to obtain cDNA populations enriched in 
cDNAs derived fixnn transcripts that are differentially expressed in colon tumor 

10 samples. By this methodology, mRNA populations were isolated firom colon tumor and 
metastatic tumor samples C'testef' mKNA) as weU as firom normal tissues, such as 
brain, pancreas, bone mairow, liver, heart, lung, stomach and smalt intesdne ("driver'* 
mKNA). From flie tester and driver mRNA populations, cDNA was synthesized by 
standard methodology. See, e.g., Ausubel, F,M. et al.. Short Protocob in Molecular 

15 j5/otogy(4^ed., John Wiley and Sons, Inc., 1999). 

The subtraction steps wctc performed using a PCR-based protocol that 
was modified to generate firagments larger than would be derived by the Clontech 
methodology. By this modified protocol, the tester and driver cDNAs were separately 
digested with five restriction endonucleases (MIu I, Msc I, Pvu H, Sal I and Stu I) each 

20 of which recognize a unique 6-base pah: nucleotide sequence. This digestion resulted in 
an average cDNA size of 600 bp, rather than Ae average size of 300 bp that results 
from digestion wiA Rsa I according to flie Clontech methodology. This modification 
did not affect the ultunate subtraction efficiency. 

Following the restriction digestion, adapter oligonucleotides having 

25 unique nucleotide sequences were ligated onto the 5' ends of the tester cDNAs; adapter 
oligonucleotides were not ligated onio the driver cDNAs. The tester and driver cDNAs 
were subsequently hybridized one to the other usmg an excess of driver cDNA. This 
hybridization step resulted in populations of (a) unhybridized tester cDNAs, (b) tester 
cDNAs hybridized to other tester cDNAs, (c) tester cDNAs hybridized to driver 

30 cDNAs, (d) unhybridized driver cDNAs and (e) driver cDNAs hybridized to driver 
cDNAs. 



BNSDOCK): <WQ ^0196388A9JL> 



wo 01/96388 



94 



PCT/US01/18S57 



Tester cDNAs hybridized to other testa: cDNAs were selectively 
amplified by a polymerase chain reaction (PGR) employing primers complementary to 
the ligated adapters. Because only tester cDNAs were lighted to adapter sequences, 
neither nnhybridized tester or driver cDNAs, tester cDNAs hybridized to driver cDNAs 

5 nor driver cDNAs hybridized to driver cDNAs were amplified using adq)ter specific 
oligonucleotides- The PGR amplified tester cDNAs were cloned into the pCR2.1 
plasmid vector (Invitrogen; Carlsbad, CA) to create libraries enriched in differentially 
expressed colon tumor antigen and colon metastatic tumor antigen specific cDNAs. 

Three thousand clones fi:Dm the pGR2.1 tumor antigen cDNA libraries 

10 were randomly selected and used to obtain clones for microarray analysis and 
nucleotide sequencing. The cDNA insert jQx>m each pGR2.1 clone was PGR amplified 
as follows. Briefly, 0.5 ^il of glycerol stock solution was added to 99,5 nl of PGR mix 
containing 80 id H2O, 10 pi lOX PGR Buffer, 6 fJ MgGli, 1 ^il 10 mM dNTPs, 1 pi 
100 mM M13 forward primer (GACGAGGTTGTAAAAGGAGGG; SEQ ID 

15 NO:2236), 1 100 mM M13 reverse primer (GAGAGGAAAGAGGTATGACG; SEQ 
ID NO:2237), and 0.5 pi 5 uAnl Taq DNA polymerase. The M13 forward and rev^ 
primers used herein were obtained firom Qperon Technologies (Alameda, CA). The 
PGR amplification was performed for thirty cycles under the following conditions: 
95°G for 5 minutes, 92'*G for 30 seconds, 57°G for 40 seconds, IS^'C for 2 minutes and 

20 75**G for 5 minutes. 

mRNA expression levels for representative clones were detmnined 
using microarray technology in colon tumor tissues (n=25), normal colon tissues (n=6), 
kidney, lung, liver, brain, heart, esophagus, small intestine, stomach, pancreas, adrenal 
gland, salivary gland, resting PBMG, activated PBMG, bone marrow, dendritic cells, 

25 spinal cord, blood vessels, skeletal muscle, skin, breast and fetal tissues. An exemplary 
methodology for performing the microarray analysis is described in Schena et al. 
Science 270:467-470. The number of tissue samples tested in each case was one (n=l), 
except where specifically noted above; additionally, all the above-motioned tissues 
were derived from humans. 

30 The PGR amplification products were dotted onto slides in an array 

format, with each product occupying a unique location in the array. mRNA was 



BNSDCXID: <WO ^0196388A2_L> 



wo 01/96388 



95 



PCT/OS01/18S57 



extracted firom the tissue sample to be tested, and fhioiescent-labeled cDNA piobes 
vfext generated by reverse transcription, according to standard methodology, in &e 
presence of fluorescmt nucleotides \|^5 and \|f3. See, e.g., Ausubel, et al., sigfra for 
exemplary reaction conditions for performing flie reverse transcription reaction; \|f S and 

5 \\f3 fluorescent labeled nucleotides may be obtained, e.g.> firom.Amersham Phannacia 
(Uppsala, Sweden) or NEN® Life Science Products, Inc. (Boston, MA). The 
microarrays were probed with the fluorescent-labeled cDNAs, flie slides were scanned 
and fluorescence intensity was measured. Genetic MicroSystems instrumentation for 
preparing the cDNA microarrays and for measuring fluorescence intensity is available 

10 ftom Aflfymetrix (Santa Clara, CA). 

An elevated fluorescence intensity in a microarray sector probed with 
cDNA probes obtained from a colon tumor or colon metastatic tumor tissue as 
compared to the fluorescence intensity in the same sector probed with cDNA piobes 
obtamed from a normal tissue mdicates a tumor antigen gene tiiat is differentially 

1 S ^pressed in colon tumor or colon metastatic tumor tissue. 

Clones disclosed herein as SEQ ID NO:l-2231 were identified from the 
PCR subtracted differential colon tumor and colon metastatic tumor cDNA libraries by 
the microairay based methodology. 

20 EXAMPLE 2 

Full-Length sequence of C93 IP Colon Tumor Protein cDNA 

This example discloses the fiill-lengfli sequence of the C931P colon 
tumor protein cDNA by LifeSeq Gold Datamining. 

The original sequence for C931P (disclosed herein as SEQ ID NO:lg61) 
25 was used as a query sequence in a BlasfN search of flie LifeSeq Gold database 
(December 2000 release). C931P matched a single LifeSeq Gold gene bm (#4751 13) 
that contained 5 template sequences. The 5 template sequences were aligned with 
C931P in order to determine a consensus, fulUengih cDNA sequence for C931P that 
encodes a 371 amino acid ORF (SEQ ID NO:2232 and 2235, respectively). Multiple 
30 splice variants were discovered in the LifeSeq Gold database. Two of these splice 



BNSOOCtO: <WO ^0196388A^JL> 



wo 01/96388 



96 



PCTAJSOl/18557 



variants are disclosed herein, one of which encodes the 371 amino add open reading 
frame. 

The original clone isolated for C931P (443 bp) was used as a query 
sequence in a BlastN search of the LifeSeq Gold ''LOtem^atesSep20(X)'' search 
S database. This search was done using the LifeSeq gold Web interface provided by 
Incyte. There was an identical match to a single LifeSeq Gold template sequence, 
number 475113.7. Then, information regarding gene bin 475113 (to which the 
475113.7 sequence belongs) was obtained from the LifeSeq Gold database. This 
475113 gene bin consisted of 5 template sequences and 176 clones. The 5 template 

10 sequences were aligned with the C93 IP sequence using the DNAStar Seqman program. 
Alignment of these sequences showed that each of the 5 template sequences represented 
an alternative splice form-each sequence had a unique multiple base pair deletion 
relative to the other sequences. This multiple sequence alignment was used to derive a 
single sequence that should rq}resent the mature mRNA for the C931P gene, as well as 

IS a single open reading jB:ame. This predicted mature mRNA sequence was obtained by 
incorporating all multiple base pair deletions that were present in the 5 template 
sequoices. The C931P original isolate sequence corresponds to a portion of the gene 
that is deleted in the predicted fully processed mRNA sequence. Thus, herein 3 
sequences obtained from LifeSeq are disclosed. One corresponds to a predicted 

20 partially spliced form of Ae gene (SEQ ID NO:2233) that wiU align with the C931P 
original isolate sequence. The second corresponds to the predicted, ftdly processed 
mRNA sequence (SEQ ID NO:2232) that will not align with the C931P original isolate 
sequence. The third sequence corresponds to the predicted coding portion of the 
sequence only (SEQ ID NO:2234). A single protein sequence is dislosed herein— the 

25 predicted C931P full-length protein sequence (SEQ ID NO:2235). 

EXAMPLE3 

MRNA EXPRESSION ANALYSIS OF THE C931P COLON TUMOR ANTIGEN USING REAL-TIME 

PGR 

30 The colon tumor candidate gene C931P (full loigth cDNA set forth in 

SEQ ID NO:2232) was analyzed by real-time PGR, as described below, using the short 



BNSDOaD: <WO ^0196388AaL,L> 



wo 01/96388 



97 



PCTAISOl/18557 



and extended colon panel. This gene vm found to have increased expression in about 
50% of colon tumors. Some expression was also observed in lymph nodes and thymus. 

The first-strand cDNA to be used in the quantitative real-time PGR was 
synthesized from 2Q\ig of total RNA that had been treated with DNase I (Amplification 

5 Grade, Gibco BRL Life Technology, Gaitherburg, MD), usmg Superscript Reverse 
Transcriptase (RT) (Gibco BRL Life Technology, Gaitherburg, MD). Real-time ?CR 
was performed with a GeneAmp™ 5700 sequence detection system (PE Biosystems, 
Foster City, CA). The 5700 system uses SYBR™ green, a fluorescent dye that only 
intercalates into double stranded DNA, and a set of gene-specific forward and reverse 

10 primm. The increase in fluorescence is monitored during the whole amplification 
process. The optimal concentration of primers was determined using a checkerboard 
approach and a pool of cDNAs fix>m breast tumors was used in fliis process. 

The PGR reaction was p^ormed in 25(il volumes that include 2.5fd of 
SYBR green bufifer, 2iil of cDNA template and 2.5fxl each of the forward and reverse 

15 primers for the gene of intl^rest. The cDNAs used for RT reactions were diluted 1:10 
for each gene of interest and 1:100 for the P-actin cottitrol. In order to quantitate the 
amount of specific cDNA (and hence initial mRNA) m the sample, a standard curve is 
generated for each run using the plasmid DNA containing the gene of interest. 
Standard curves were generated using the Ct values detemiined in the real-time PGR 

20 which were related to the initial cDNA concentration used in the assay. Standard 
dilution ranging firom 20-2x10^ copies of the gene of interest was used for this purpose. 
In addition, a standard curve was generated for p-actin ranging from 200fg-2000fg. 
This enabled standardization of the initial RNA content of a tissue sample to the amount 
of p-actin for comparison purposes. The mean copy number for each group of tissues 

25 tested was normalized to a constant amount of p-actin, allowmg the evaluation of the 
over-expression levels seen with each of the genes. 



BNSDOCIO: <WO__OI9e38BA?„l^ 



WO01^M»388 



98 



PCT/US01/18S57 



EXAMPLE4 
Peptide Priming Of T-helper Lines 
Generation of CD4^ T helper lines and identification of peptide q>itopes 
derived from tumor-specific antigens that are capable of beii^ recognized by CD4* T 
5 cells in the context of HLA class II molecules, is carried out as follows: 

Fifteen-mer peptides overlapping by 10 amino acids, derived fix)m a 
tumor-specific antigen, are generated using standard procedures. Dendritic cells (DC) 
are derived fit)m PBMC of a normal donor using GM-CSF and IL-4 by standard 
protocols. CD4^ T cells are generated from tfie same donor as the DC using MACS 
10 beads (Miltenyi Biotec, Auburn, CA) and negative selection. DC are pulsed overnight 
with pools of the 15-mer peptides, with each peptide at a final concentration of 0.25 
Hg/ml. Pulsed DC are washed and plated at 1 x 10"* cells/well of 96-well V-bottom 
plates and purified CD4* T cells are added at 1 x 10^/welL Cultures are supplemented 
with 60 ng/ml IL-6 and 10 ng/ml IL-12 and incubated at 37°C. Cuhures are 
IS restimulated as above on a weekly basis using DC generated and pulsed as above as 
antigen presenting cells, supplemented with 5 ng/ml IL-7 and 10 U/ml IL-2. FoUowmg 
4 in vitro stimulation cycles, resulting CD4* T cell lines (each line corresponding to one 
well) aie tested for specific proliferation and cytokine production in response to the 
stimulating pools of peptide with an irrelevant pool of peptides used as a controL 

20 

EXAMPLES 

Generation of Tumor-Speofic CTL Lines Using In Vitro Whole-Gene Priming 

Using in vitro whole-gene priming with tumor antigen-vaccinia infected 
DC (see* for example, Yee et al. The Journal of Immunology, 157(9):4079-86, 1996), 

25 human CTL lines are derived that specifically recognize autologous fibroblasts 
transduced with a specific tumor antigen, as determined by interferon-y ELISPOT 
analysis. Specifically, dendritic cells pC) are differentiated fi-om monocyte cultures 
derived firom PBMC of normal human donors by growing for five days in RPMI 
medium containmg 10% human serum, 50 ng/nd human GM-CSF and 30 ng/ml human 

30 IL-4. Following culture, DC are infected overnight with tumor antigen-recombinant . 
vaccmia virus at a muhiplicity of infection (M.O.I) of five, and matured overnight by 



BNSOOCtD: <W0 ^OI9638aASJL^ 



wo 01/96388 PCTA)S01/18557 

99 

tiie addition of 3 ^g^ml CD40 ligand. Virus is then inactivated by UV irradiation. 
CD8+ T cells are isolated using a magnetic bead system, and priming cultures are 
initiated using standard culture techniques. Cultures are restimulated every 7-10 days 
using autologous primary fibroblasts retrovirally transduced with,previously identified 
5 tumor antigens. Following four stimulation cycles, CD8+ T cell lines are identified that 
specifically produce interferon-y when stimulated with tumor antigen-transduced 
autologous fibroblasts. Using a panel of HLA-mismatched B-LCL lines transduced 
with a vector expressing a tumor antigen, and measuring interferon-y production by the 
CTL lines in an ELISPOT assay, the HLA restriction of the CTL lines is determined. 

10 

EXAMPLE 6 

Generation and Characterizatign of anti-Tumor Antigen monoclonal 

antibodies 

Mouse monoclonal antibodies are raised against £ coli derived tumor 
IS antigen proteins as follows: Mice are inrninnized with Complete Freund's Adjuvant 
(CPA) containing 50 ^g recombinant tumor protein, followed by a subsequent 
intraperitoneal boost with Incomplete Freund*s Adjuvant (IF A) containing lOjig 
recombinant protein. Three days prior to removal of the spleens, the mice are 
immxmized intravmously with approximately SOjig of soluble recombinant protein. 
20 The spleen of a mouse with a positive titer to the tumor antigen is removed, and a 
single-cell suspension made and used for fusion to SP2/0 myeloma cells to generate B 
cell hybridomas. The supematants from the hybrid clones are tested by ELISA for 
specificity to recombinant tumor protein, and epitope mapped using peptides that 
spanned tiie entire tumor protein sequence. The mAbs are also tested by flow 
25 cytometry for tbdr ability to detect tumor protein on the surface of cells stably 
transfected with the cDNA encoding the tumor protein. 

EXAMPLE? 
Synthesis of Polypeptides 
30 This Example discloses ah exemplary methodology for tiie preparation 

of colon tumor proteins. 



BNSDOCIO: <WO ^0196388A^J^ 



wo 01/96388 PCTAJSOl/1^7 

100 

Polypeptides may be synAesized on a Perkin Elmer/Applied Biosystems 
Division 430A peptide synthesizer using FMOC chemistry with HPTU (0- 
Benzotiiazole-N,N J^,N -tetramethyluronium hexafluorophosphate) activation. A Gly- 
Cys-Gly sequence may be attached to the amino tenninus of the peptide to provide a 

5 method of conjugation, binding to an immobilized surface, or labeling of the peptide. 
Cleavage of the peptides from the solid support may be carried out using the following 
cleavage mixture: trifluoroacetic acid:ethanedithiol:thioanisole:water:phenol 
(40:1:2:2:3). After cleaving for 2 hours, the peptides may be precipitated in cold 
methyl-t-butyl-ether. The peptide pellets may then be dissolved in water containing 

10 0.1% trifluoroacetic acid (TFA) and lyophilized prior to purification by CI 8 reverse 
phase HPLC. A gradient of 0%-60% acetonitrile (containing 0.1% TFA) in water 
(containing 0.1% TFA) may be used to elute the peptides. Following lyophilization of 
the pure fractions, the peptides may be characterized using electrospray or other types 
of mass spectrometry and by amino acid analysis. 

15 

From tiie foregoing it will be appreciated that, although specific 
embodiments of the invention have been described herein for piuposes of illustration, 
various modifications may be made without deviating from the spirit and scope of the 
invention. Accordingly, the invGotion is not limited except as by the appended claims. 



BNSDOCID: <WO__019e388A?JL> 



wo 01/96388 PCT/USOl/18557 

101 

CLAIMS 

What is claimed is: 



1. An isolated polynucleotide comprising a seguence selected from 
the group consisting of: 

(a) sequences provided in SEQ ID NO:l-2234; 

(b) complements of the sequences provided in SEQ ID NO: 1-2234; 

(c) sequences consisting of at least 20 contiguous residues of a 
sequence provided in SEQ ID NO:l-2234; 

(d) sequences that hybridize to a sequence provided in SEQ ID ' 
NO: 1 -2234, under moderately stringent conditions; 

(e) sequences having at least 75% identity to a sequence of SEQ ID 

NO:I-2234; 

(f) sequences having at least 90% identity to a sequence of SEQ ID 
NO:l-2234;and 

(g) degenerate variants of a sequence provided in SEQ ID 

NO:l-2234. 



2. An isolated polypeptide comprising an amino acid sequence 
selected from the group consisting of: 

(a) sequences encoded by a polynucleotide of claim 1 ; and 

(b) sequences having at least 70% identity to a sequence encoded by 
a polynucleotide of claim 1; and 

(c) sequences having at least 90% identity to a sequence encoded by 
a polynucleotide of claim 1; and 

(d) sequences provided in SEQ ID NO:2235. 

3. An expression vector comprising a polynucleotide of claim 1 
operably linked to an expression control sequence. 



l9638aA9JL;> 



wo 01/96388 



102 



PCTAJSOl/18557 



4. A host cell transformed or transfected with an expression vector 
according to claim 3. 

5. An isolated antibody, or antigen-bindmg fiagment thereof, that 
specifically binds to a polypeptide of claim 2. 

6. A method for detecting the presence of a cancer in a patient, 
comprising the steps of: 

(a) obtaining a biological sample from the patient; 

(b) contacting the biological sample with a binding agent that binds 
to a polypeptide of claim 2; 

(c) detecting m the sample an amoimt of polypeptide that binds to 
the binding agent; and 

(d) comparing the amount of polypeptide to a predetermined cut-off 
value and therefrom determining the presence of a cancer in the patient 

7. A fusion protein comprismg at least one polypeptide according to 

claim 2. 

8. An oligonucleotide that hybridizes to a sequence recited in SEQ . 
ID NO:l-2234 under moderately stringent conditions. 

9. A method for stimulating and/or expanding T cells specific for a 
tumor protein, comprismg contacting T cells with at least one component selected fix»m 
the group consisting of: 

(a) polypeptides according to claim 2; 

(b) polynucleotides according to claim 1 ; and 

(c) antigen-presendng cells that express a polynucleotide according 

to claim 1, 

under conditions and for a time sufGcient to pennit the stimulation 
and/or e)q)ansion of T cells. 



BNSDOOD: «WO ^D196388A^^ 



wo 01/96388 



103 



PCT/DS01/18S57 



10. An isolated T cell population, comprising T cells prepared x 
according to the method of claim 9. 

11. A composition comprising a first component selected fipom the 
group consisting of physiologically acceptable carriers and inunmiostimulants, and a 
second component selected from the group consisting of: 

(a) polypeptides according to claim 2; 

(b) polynucleotides according to claim 1 ; 

(c) antibodies according to claim 5; 

(d) fusion proteins according to claim 7; 

(e) T cell populations according to claim 10; and 

(f) antigen presenting cells that exptess a polypeptide according to 

claim 2. 

12. A method for stimulating an immune response in a patient, 
comprising administering to the patient a coniposition of claim 1 1 . 

13. A method for the treatment of a cancer in a patient, comprising 
administering to the patient a composition of claim 11. 

14. A method for determining the presence of a cancer in a patient, 
comprising the steps of: 

(a) ' obtaining a biological sample from the patient; 

(b) contacting the biological sample mth an oligonucleotide 
according to claim 8; 

(c) detecting in the sample an amount of a polynucleotide that 
hybridizes to the oligonucleotide; and 

(d) compare the amount of polynucleotide that hybridizes to the 
oligonucleotide to a predetem[iin6d cut-off value, and therefrom determining the 
presence of the cancer in the patient. 



BNSDOCO): <W0 0196388MLL^ 



wo 01/96388 



104 



PCTAJSD1/185S7 



15. A diagnostic kit comprising at least one oligonucleotide 
according to claim 8. 

16. A dis^ostic kit comprising at least one antibody according to 
claim 5 and a detection reagent, wherein the detection reagent comprises a reporter 
group. 

1 7. A method for inhibiting the development of a cancer in a patient, 
comprising the steps of: 

(a) mcubating CD4+ and/or CD8+ T cells isolated fiom a patient 
with at least one component selected fiom the group consisting of: (i) polypeptides 
according to claim 2; (ii) polynucleotides according to claim 1; and (iii) antigen 
presenting cells that express a polypeptide of claim 2, such that T cell {proliferate; 

(b) administering to the patient an effective amount of the 
proliferated T cells, 

and thereby inhibiting the development of a cancer in the patient 



BNSDOCtD: <WO__OI86388A2_I^ 



(12) INTERNATIONAL APPUCATION FUBUSHED UNDER THE PATENT COOPERAIKM TREATY (PCT) 



CORRECTED VERSION 



(ISO WDrld Intellectual Property Orgintxa.tim 
Inteniatioiial Bureau 

(43) Interaational Publication Date 
20 December 2001 (20.12J1001) 




PCT 



inuiiiiiiiiiiiiiiiiiii 

(10) iDternational Publication Number 

WO^l/96388 A2 



(51) International Patent Classification^: C07K 14/47 

(21) InternatbnalAppUcation Number: FCTAJSOl/lRSS? 

(22) Inteniatronal Filing Date: 8 Jcme 2001 (08.06.2001) 



(25) FiUng Language: 

(26) Publication Language: 



EngUsb 
English 



(30) Priority Data: 
6(y270^16 



9 June 2000 (09.06.2000) US 
20 Fdmiaiy 2001 (20.012001) US 



(71) Applicant ffor all designated States except US}: CORIXA 
CORPORATION [US/US]; Suite 200. 1124 Cohimbia 
Street. Seattle, WA 98104 (US). 

(72) Inventors; and 

(75) Inventors/Applicants ffor US only): JUNG, Yuqiu 
[CNAJS); 5001 S. 232nd Street, Kent. WA 98032 (US). 
HARLOCKER, Susan, L. [US/US); 7522 13th Av- 
enue W., SeatUe, WA 98117 (US). S£CRIST» Heather 
[US/US]: 3844 35th Avenue W., Seatde, WA 98199 (US). 

(74) Agents: POTTER, Jane, R^ Seed Intellecmal Plop- 
eny Law Group PLLC Suite 6300, 701 Hfih Avenue. Seat- 
tle, WA 98104-7092 et (US). 

(81) Designated States (national): AE. AG. AL, AM. AT. AU. 
AZ. BA, BB, BG, BR. BY, BZ, OA, CH, CN, CO. CR, CU, 



CZ. DE, DK. DM. DZ. EC. EE, ES. H. GB. GD. GE. GH. 
GM, HR, HU. H). IL. IN. IS, JP, KE. KG. KP. KR, KZ. LC. 
LK. LR, LS. LT. LU. LV, MA, MD. MG. MK, MN, MW. 
MX. MZ. NO. NZ. PL, FT. RO. RU, SD, SE. SG. SI, SK, 
SL. TJ. TM. TR. TT. UA, UG, US, UZ, VN, YU. ZA. 
ZW 

(84) Designated States (regional): ARIPO patent (GH. GM, 
KE, LS. MW. MZ. SD. SL. SZ. TZ. UG, ZW). Eurasian 
patent (AM . AZ. B Y. KG. KZ. MD. RU. TJ. TM). European 
patent (AT. BE. CH. CY, DE. DK. ES. H. FR. GB, GR, IE. 
IT, LU. MC. NL, FT. SE, TR), OAPI patent (BF, BJ, CF, 
CG, CI, 04. G A. GN. GW, ML, MR. NE, SN, TD. TG). 

Published: 

— without international search report and to be republished 
upon receipt of that report 

— with sequence listing part of description published sepct- 
ratefy in electronic form and available iq>on request from 
the International Btaeau 

(48) Date of publication of this corrected version: 

21 March 2002 

(15) Information about Correction: 

see PCT Gazette No. 12/2002 of 21 Maich 2002. Section 

n 

For two-letter codes and other abbreviations, refer to the '^Guid- 
ance Notes on Codes and Abbreviations" of^aring at the begin- 
ning of eadi regular issue of the PCT Gazette. 



< 
do 

00— — 

(54) title: compositions and methods for the therapy and diagnosis of colon cancer 

On 

^ (57) Abstract: Compositions and methods for the ther^y and diagnosis of cancer, such as colon cancer, are disclosed. Composi- 
^ tions may comprise one or more colon tumor proteins, immunogenic portions thereof, or polynucleotides that encode such portions. 

Alternatively, a therapeutic cotoposition may comi^ise an antigen presenting cell that expresses a colon tumor protein, or a T cell 
O that is specific for ceUs expressing such a protein. Such compositions may be used, for example, for the prevention and treatment of 
^ diseases such as cdon canceL Diagnostic mediods based on detecting a colon tumor piotem. or mRNA encodii^ such a protdn, in 
^ a sample are also provided. 



*iWO qi96a88A2.IA> 



(12) INTERNATIONAL APPUCATION PUBLISHED UNDER THE PATENT COOPERATION TREATY (PCT) 



(19) World Intellectual Property Organization 
International Bureau 

(43) Internatiooal Publication Date 
20 December 2001 (20.12.2001) 




PCT 



liiiiiiiiiiiiiiiiiiigniiie 

(10) iDteniational Pnbttcation Number 

WO 01/096388 A3 



(51) IntematioBal Patent Classificatloii^: C07K 14/47. 
14/82, C12N lS/12, 15/62 

(21) lotematiooal Application Number: PCT/US01/18S57 

(22) iDtematiODal Fillog Date: 8 June 2001 (08.06.2001) 



(25) Filing Language: 

(26) Publication Language: 



English 
English 



(30) Priority Data: 

60/210.899 
60/270^16 



9 June 2000 (09.06.2000) US 
20 Februaiy 2001 (20.02.2001) US 



(71) Applicant (/Tor all designated States except US): CORIXA 
CORPORATION [US/US]; Suite 200. 1124 Columbia 
Street. Seattle, WA 98104 (US). 

(72) Inventors; and 

(75) Inventors/Applicants ffbr US ortfy): JIANG, Yuqiu 
[CN/US]; 5001 S. 232nd Street, Kent, WA 98032 (US). 
HARLOCKER, Susan, L. [US/US]; 7522 13th Av- 
enne W.. Seattle. WA 98117 (US). SECRIST, Heather 
[US/US]; 3844 35th Avenue W.. Seattle. WA 98199 (US). 

(74) Agents: POTTER, Jane^ R.; Seed Intellectual Prop- 
erty Law QroupPLLC, Soite 6300, 701 Rfth Avenue, Seat- 
Ue, WA 98104-7092 et a). (US). 

(81) Designated States (national): AE. AG. AL. AM, AT, AU, 
AZ, BA. BB, BG. BR, BY, BZ. CA. CH, CN. CO, CR, CU, 



CZ. DE, DK. DM, DZJBC, EE. ES. H. GB, GD, GE. GH. 
GM, HR, HU, ID, IL, IN, IS. JP. KE, KG. KP, KR, KZ. LC, 
LK. LR. LS. LT. LU, LV, MA, MD. MG, MK, MN, MW, 
MX, MZ. NO, NZ. PL. FT. RO. RU. SD, SE. SG, SI. SK, 
SL. TJ, TM. TR, XT, TZ, UA. UG, US. UZ. VN. YU. ZA, 
ZW. 

(84) Designated States (regional): ARIPO patent (GH. GM. 
KE. LS. MW, MZ, SD, SL, SZ, TZ, UG. ZW), Eurasian 
patent (AM. AZ, BY, KG, KZ, MD. RU, TJ, TM), European 
patent (AT. BE, CH, CY. DE. DK. ES, H, FR. GB, GR. IE. 
IT, LU. MC, NL. PT. SB. TR). OAH patent (BF. BJ, CF, 
CO, CI. CM. G A. ON , GW, ML, MR, NE, SN. TD, TO). 

Published: 

— with international search report 

— with sequence listing part of description published sepa- 
rately in electronic form and available upon request from 
the International Bureau 

(88) Date of publication of the international search report: 

3October20Q2 

(15) Information about Correction: 
Previous Correction: 

see per Gazette No. 12/2002 of 21 March 2002, Section 

n 

For two4eUer codes and other abbreviations, refer to the "Guid- 
ance Notes on Codes and Abbreviations** appearing at the begin- 
ning of each regular issue of the PCT Gazette. 



00 

^ 

VO (54) Title: COMPOSmONS AND METHODS FOR THE THERAPY AND DL\GNOSIS OF COLON CANCER 
ON 

^ (57) Abstract: Compositions and methods for the therapy and diagnosis of cancer, such as colon cancer, are disclosed. Composi- 

^ tions may comprise one or more colon mmor proteins, immunogenic portions thereof, or polynucleotides that encode such portions. 
Alternatively, a therapeutic composition may compise an antigen presenting cell that expresses a colon tumor protein, or a T cell 
that is specific for cells expressing such a protein. Such compositions may be used, for examine, for the prevention and treatment of 

^ diseases such as colon cancer. Diagnostic methods based on detecting a colon tumor protein, or mRNA encoding such a protein, in 

\^ a sample are also provided. 



BNSD0C10:<W0_ 



_01963eaA3_l_> 



INTERNATIONAL SEARCH REPORT 



tlonal Application No 

PCT/US 01/18557 



A. CLASSIRCATION OF SUBJECT iMAmR , . ^ 

IPC 7 C07K14/47 C07K14/82 C12N15/12 C12N15/62 



Aocoiding lo International Patent Ctasslttcatton (IPC) orlo twlti national ciasslBcallon and IPC 



a RELDS SEARCHED 



Mmimuindocurnentalion searched (i 

IPC 7 C07K C12N 



[dasslficaUon eystem toUowed by dassHlcaiion symliols) 



DocumenlaUoR searched other than rninlmumdocumentatl^ in me fieUs searohed 



Electronic data base consulted during the Inlemationa) search (name of data t>ase and wheie practical, search tenns used) 

EPO-Internal , ENBL, UPI Data, PAJ 



C. D0CUMEN1S CONSIDERED TO BE RELEVANT 



Categny* Cnatkm of document, iMIh kidtoaltoiii whOTappapiiate.alliwnl«v^ 



RalwMilocWniNa 



SIMON B ET AL: "EPITHELIAL GLYCOPROTEIN 
IS A MEMBER OF A FAMILY OF EPITHELIAL CELL 
SURFACE ANTIGENS HOMOLOGOUS TO NIDOGEN. A 
MATRIX ADHESION PROTEIN" 
PROCEEDINGS OF THE NATIONAL ACADEMY OF 
SCIENCES OF USA, NATIONAL ACADEMY OF 
SCIENCE. WASHINGTON. US, 
vol. 87, no. 7, 1 April 1990 (1990-04-01), 
pages 2755-2759, XP000651084 
ISSN: 0027-8424 
figure 3 

-/- 



I- 9, 

II- 17 



m 



further documents are lisled In the continuation of box C. 



Patent family meml^ers are nsied In annex. 



* Special categories of died documents : 

*A' document deflnbig the general state ol the an which Is not 

considered to be of particular relevance 
*E* earlier document but published on or after the international 

QHngdale 

V document which may throw doubts on priorliy dalmCs) or 
which Is cited to establish the publicaiion date d another 
citation or other special reason (as specified) 

'O* document referrtng to an oral disdosuie. use, exhibUon or 
other means 

'P* document published prior to Ihe Iniernattonal filing dale but 
later than the priority dale dabned 



*T later document published alter the Mernational liing dale 
or priority date and not m conffict with the appDcalion but 
cited to understand the prlndpte or ihaoiy underlying Ihe 



'X* document of particular relevance: the claimed Invention 
cannot be oortsidered novel or cannot be considered to 
Involve an Inventive slep when the document b taken alone 

*V document of particular relevance; the datnwd Invention 
cannot be considered to involve an tnventtve step when the 
document is combined with one or more other such docu> 
menls. such combination being oinrious 10 a person sMIIed 
fritheart. 

*&* document member of the same patent family 



Dale Of the actual oompletion of the Iniefflalional search 



7 January 2002 



Dale of mailing of the InieFnatlonat search report 



and mailing address of the ISA 

European Patent Office. P.a 5616 Palentlaan 2 

NL-2260HVRijswQK 

Tel (+31-70) 340-2040, Tit. 31 651 epo nl 

Fax: («31-70) 340-3016 



Aulhorliedottloer 



Grosskopf, R 



Foim PCT/ISV210 ^aeond shael) (Ju^r 
BNSCXX:iD: <WO 019638aA3jL> 



page 1 of 2 



INTERNATIONAL SEARCH REPORT 



»nal Application No 

PCT/US 01/18557 



C.(CanUnuallon) DOCUMENTS CONSIOERED TO BE RELEVANT 



Caiegoiy* 



Clatbn 01 docinieni, iwU) lndicatk)n.>iiilieieapprapiiale^ 



RelevanlloclaimNo. 



PEREZ M S ET AL: "ISOLATION AND 
CHARACTERIZATION OF A CONA ENCODING THE 
KSl/4 EPITHELIAL CARCINOMA MARKER" 
JOURNAL OF IMMUNOLOGY, THE WILLIAMS AND 
UILKINS CO. BALTIMORE. US, 
vol. 142, no. 10, 

15 May 1989 (1989-05-15), pages 3662-3667, 
XP000651096 
ISSN: 0022-1767 
figure 2 

EP 0 326 423 A (LILLY CO ELI) 
2 August 1989 (1989-08-02) 
page 9 

page 16, line 27 -page 17, line 18 

SZALA S ET AL: "HOLECUUR CLONING OF CDMA 
FOR THE CARCINOMA-ASSOCIATED ANTIGEN GA 
733-2" 

PROCEEDINGS OF THE NATIONAL ACADEMY OF 

SCIENCES OF USA, NATIONAL ACADEMY OF 

SCIENCE. WASHINGTON. US, 

vol. 87, no. 9. 1 May 1990 (1990-05-01), 

pages 3542-3546, XP000566331 

ISSN: 0027-8424 

see the whole document 



I- 9. 

II- 17 



I- 9. 

II- 17 



I- 9. 

II- 17 



FoBn PCTASM210 (eenUmiBlion of Moond ahaatXJuly 1862) 
BNSDOCIO: -dWO ^0ig63eaA3LI.> 



page 2 of 2 



INTERNATIONAL SEARCH REPORT 



IKemationalappllcalion No. 
PCT/US 01/18557 



Box I Observations where certain claims were found unsearcfiable (Continuation of item 1 of first siieet) 



This international Search Report has not been estal)Ushed in respect of certain daims under Article 17(2)(a} lor the following reasons: 
1. [71 Claims Nos^- 

because they relate to subject matter not required to be searched by this Authority, namely: 

Although claims 9, 12, 13 and 17 are directed to a method of treatment of the 
human/animal body, the search has been carried out and based on the alleged 
effects of the compound/composition. 



Claims Nos.: " 

because they relate to parts of the Internattonal Application that do not comply with the prescribed requirements to such 
an extent that no meaningM international Search can be carried out, spedficalty: 

see FURTHER INFORHATION sheet PCT/ISA/210 



3. J_| Claims Nosj 

— because they are dependent claims and are not drafted in accordance with the second and tNrd sentences of Rule 6.4(a). 



Box H Observations where unity of invention is laclcing (Continuation of item 2 of first sheet) 



This Intemalional Searching Authority found multiple inventions in this imemaHonal application, as follows: 



see additional sheet 



1. I I As all required additional search feeswere timely paid by ttte applicant, this Intsmatkmal Search Report covers bU 
' — * searchable claims. 

2. Q As all searchable claims could be searched without effort justifying an addltional fee. this Authortfy did not invite payment 

of any additional lee. 



3. I I As only some of the required additional search fees were timely paid by the applicant, this Intemalional Search Report 
■ — ' covers only those claims lor which fees were paid, specifically claims Nos.: 



4. [yj No required addnional search fees were timely paid by the applicant. Consequently, this international Search Report is 
restricted Id the invention first mentioned \n ttie claims; It Is covered by claims Nos.: 

1-9, 11-17 (all partially) 



Remark on Protest 



[ I The additional search fees were accompanied by the applicant's protest 
[ [ No protest accompanied the payment of additional search lees. 



farm PCT/lSA/210 (continuation of first sheet (i)) (July 1898) 

BNSDOCID: <WO__pi06ae8A3LL> 



Iniemattonal Appiication No. PCT/US 01 A8557 

FURTHER INFORMATION CONTINUED FROM PCTASA^ 210 

This International Searching Authority found multiple (groups of) 
Inventions In this International application, as follows: 

1. Claims: 1-9, 11-17 (all partially) 

The claims Insofar as they relate to SEQ ID- NO: 1 

2. Claims: 1-9, 11-17 (all partially) 

Inventions 2 to 2235; Claims Insofar as they relate to SEQ 
ID NO:s 2 to 2235 

3. Claims: 10 and 11-13 (partially) 

Invention 2236: A T-cell population and the use thereof 



BNSOOaO*. <WO__^019e38aA3_L> 



tmemationalApplfcatlonNO. PCTAlS 01 il8557 



FURTHER INFORMATION CONTINUED FROM PCT/ISA/ 210 



Continuation of Box 1.2 



Items (d) to (g) of Claim 1 are directed to a multitude of undefined 
nucleotide fragments which are neither defined by their function nor by 
their length. Thus, the scope of Claim 1 In this resepct (and 
consequently the scope of all other claims relating to these items) Is 
totally unclear and unlimited and renders a meaningful or complete 
research Impossible. 

The same applies for the oligonucleotides according to Claim 8 (and 
consequently the method and the diagnostic kit according to Claims 14 
and 15). 

The applicant's attention is drawn to the fact that claims, or parts of 
claims, relating to Inventions in respect of which no International 
search report has been established need not be the subject of an 
international preliminary examination (Rule 66.1(e) PCT). The applicant 
is advised that the EPD policy when acting as an International 
Preliminary Examining Authority is normally not to carry out a 
preliminary examination on matter which has not been searched. This is 
the case irrespective of whether or not the claims are amended following 
receipt of the search report or during any Chapter II procedure. 



BNSOOCID: <WO__019e38aA3JL> 



INTERNATIONAL SEARCH REPORT 

liilciii M t l ewoiip»l«nttiiiill|rw«inb«i» 



InVbhrnBl AppUcollM No 

PCT/US 01/18557 



Patent document 
cited in search report 


PubBcaMon 
date 


Patent family 
memt)er(s) 


Publication 
date 


EP 0326423 A 


02-08-1989 


CA 


1340221 Al 


15-12-1998 






DE 


68922757 Dl 


29-06-1995 






DE 


68922757 T2 


16-11-1995 






DK 


35189 A 


31-07-1989 






EP 


0326423 A2 


02-08-1989 






JP 


2005867 A 


10-01-1990 






JP 


2774298 B2 


09-07-1998 






US 


5348887 A 


20-09-1994 



Foiin PCTJISAtttO <peieni fami^ mwm) <JiiV tSSe) 



BNSDOCID: <W0 919638BA3^L> 



This Page is Inserted by IFW Indexing and Scanning 
Operations and is not part of the Official Record 

BEST AVAILABLE IMAGES 

Defective images within this document are accurate representations of the original 
documents submitted by the appHcant. 

Defects in the images include but are not limited to the items checked: 

□ BLACK BORDERS 

□ IMAGE CUT OFF AT TOP, BOTTOM OR SffiES 

□ FADED TEXT OR DRAWING 

□ BLURRED OR ILLEGIBLE TEXT OR DRAWING 

□ SKEWED/SLANTED IMAGES 

□ COLOR OR BLACK AND WHITE PHOTOGRAPHS 

□ GRAY SCALE DOCUMENTS 

□ LINES OR MARKS ON ORIGINAL DOCUMENT 

□ REFERENCE(S) OR EXHIBIT(S) SUBMITTED ARE POOR QUALITY 

□ OTHER: 

IMAGES ARE BEST AVAILABLE COPY. 
As rescanning these documents will not correct the image 
problems checked, please do not report these problems to 
the IFW Image Problem Mailbox. 



